ROSS: Relearning from Self-Generated Rollouts through Selective Supervision Paper • 2609.35954 • Published 6 days ago • 46
OmniTaskonomy: When Does Visual Generation Improve Visual Understanding? Paper • 2609.38079 • Published 5 days ago • 52
Surprising Success, Repeated Failure: Entropy-Guided Credit Assignment for Exploration in LLM Reasoning Paper • 2609.33781 • Published 7 days ago • 43
Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 11 days ago • 55
Periodic Weak Spots: Phase Sensitivity from Chunked KV-Cache Compression Paper • 2609.36322 • Published 6 days ago • 106
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 6 days ago • 379
Chinese-Jev: Bringing System One Model to Chinese-Language Tasks Paper • 2609.36965 • Published 5 days ago • 22
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 7 days ago • 550
Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR Paper • 2609.37868 • Published 5 days ago • 58
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 9 days ago • 159
Paragraph Boundaries Are Not White Space:Compression Depth as the Signature of Hierarchical Structure Paper • 2609.23551 • Published 14 days ago • 7
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 12 days ago • 91
TimeEvo: Failure-Driven Self-Evolution of a Time Series Agent Paper • 2609.27277 • Published 11 days ago • 32