DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 2 days ago • 67
Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation Paper • 2609.20744 • Published 2 days ago • 8
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 2 days ago • 23
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 4 days ago • 29
Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems Paper • 2609.17320 • Published 4 days ago • 2
PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control Paper • 2609.17521 • Published 4 days ago • 6
SNAP3D: Physically Grounded 3D Parts for Assembly from a Single Image Paper • 2609.13146 • Published 8 days ago • 3
SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking Paper • 2609.13141 • Published 8 days ago • 61
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics Paper • 2609.10712 • Published 10 days ago • 41
Recursive Code World Models: Building Complex Worlds through Recursive Scene Programs Paper • 2609.11499 • Published 9 days ago • 33
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 9 days ago • 264
AgenticGen: Reward-Guided Agentic Video Generation for Advertising Paper • 2609.09187 • Published 19 days ago • 9
Φ-Bench: Can Large Language Models Engineer the Infrastructure That Powers Them? Paper • 2609.10226 • Published 10 days ago • 19
TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model Paper • 2609.09158 • Published 11 days ago • 22
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents Paper • 2609.09153 • Published 11 days ago • 40