n1ghtf4l1/Agentic-Diagnostic-Reasoning-with-Multimodal-SLMs-via-Reinforcement-Learning Updated Nov 21, 2025 • 128 • 12
Morphometric Imitation: From Morphology and Contact Aware Hand Retargeting to Sim-to-Real Visuomotor Policy Paper • 2609.28660 • Published 11 days ago • 15
EVO-WAM: Evolving World Action Models through Video-Action Verification Paper • 2609.38057 • Published 5 days ago • 38
tomyimkc/repro-optimal-regret-for-policy-optimization-in-contextual-bandits-traces Traces • Updated Jul 26 • 1 • 72 • 2
FocusVTC: Efficient and High-Performance Visual Text Compression with Adaptive Resolution Paper • 2609.36651 • Published 5 days ago • 29
SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video Paper • 2609.37969 • Published 5 days ago • 36
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 7 days ago • 550
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 7 days ago • 43
Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL Paper • 2609.32577 • Published 8 days ago • 130
Rethinking Training-Inference Mismatch in LLM Reinforcement Learning: Where It Arises and How to Correct It Paper • 2609.32444 • Published 8 days ago • 30
Imprint Reader: From Weight-Update Readout to Behavioral Intervention Paper • 2609.35261 • Published 6 days ago • 16