Beyond the Current Scene: Event-Referential Grasping with Active View Selection Paper • 2609.39375 • Published 3 days ago • 29
ROWBench: Do Video Models Render What the Program Specifies? Paper • 2610.02205 • Published 2 days ago • 46
Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States Paper • 2610.01415 • Published 2 days ago • 69
AutoGUIWorld: Image Generators as Visual World Models for GUI Agent Paper • 2610.01215 • Published 2 days ago • 40
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF Image-Text-to-Text • 177B • Updated 3 days ago • 1.14M • 462