Where to Look Matters: On-Policy Self-Distillation for Long-Video Understanding Paper • 2608.25356 • Published 1 day ago • 20
Apodex 1.1: Scaling Agentic Intelligence for Complex Work Paper • 2608.23283 • Published 3 days ago • 188
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Paper • 2608.06867 • Published 20 days ago • 109
Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning Paper • 2608.02831 • Published 24 days ago • 14
Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning Paper • 2608.02831 • Published 24 days ago • 14
AudioRubrics Collection Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning: model and rubric dataset. • 2 items • Updated Jul 12 • 1
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Paper • 2608.06867 • Published 20 days ago • 109
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Paper • 2608.05000 • Published 21 days ago • 62
ArcMemo: Abstract Reasoning Composition with Lifelong LLM Memory Paper • 2509.04439 • Published Sep 4, 2025 • 2
TS-Reasoner: Aligning Time Series Foundation Models with LLM Reasoning Paper • 2510.03519 • Published Oct 3, 2025 • 1
FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse Paper • 2606.11290 • Published Jun 9 • 2
ArrowGEV: Grounding Events in Video via Learning the Arrow of Time Paper • 2601.06559 • Published Apr 16 • 1
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 24 days ago • 181
Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning Paper • 2608.02831 • Published 24 days ago • 14