VisionHOPE: Visual Backbones as Self-Modifying Learning Systems Paper • 2609.33325 • Published 5 days ago • 320
SLCA-GRPO: Resolving Cross-Segment Credit Misattribution in Tool-Calling RL Paper • 2609.29050 • Published 8 days ago • 13
TRACE: Temporal Audit and Condition-aware Evaluation of Streaming Video Understanding Paper • 2609.30670 • Published 7 days ago • 11
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 8 days ago • 32
ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds Paper • 2609.30199 • Published 8 days ago • 28
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 13 days ago • 41
Fingers as Legs: Learning Self-Supported Locomotion and Manipulation with an Anthropomorphic Hand Paper • 2609.17172 • Published 17 days ago • 5
GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 11 days ago • 72
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 11 days ago • 102
JEV-as-a-Judge: Accept When Confident, Escalate When Unsure Paper • 2609.26550 • Published 10 days ago • 42