Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Paper • 2607.18789 • Published 5 days ago • 1
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published 12 days ago • 108
Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments Paper • 2605.22189 • Published May 21 • 8
Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence Paper • 2605.30093 • Published May 28 • 15
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Paper • 2605.20266 • Published May 18 • 56