Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published 14 days ago • 83
Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vision-Language-Action Models Paper • 2606.19297 • Published Jun 17 • 79
RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation Paper • 2607.06558 • Published 17 days ago • 79
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 172
Running 135 Unfolding Robotics: Open-Source Shirt Folding from Data to Deployment 🤖 135 Learn to teach a robot to fold your clothes