Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds Paper • 2608.02636 • Published Jul 31
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 10 days ago • 26
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 10 days ago • 26
HPSD: Hybrid-Policy Self-Distillation for Text-Image-to-Video Diffusion Models Paper • 2608.13205 • Published about 1 month ago • 3
Intern-S2-Preview: Scientific Agentic Foundation Model Paper • 2608.13505 • Published about 1 month ago • 72
AdaGRPO: A Capability-Aware Adaptive Enhancement for Flow-based GRPO Paper • 2606.06828 • Published Jun 5
CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning Paper • 2606.09393 • Published Jun 8
Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games Paper • 2606.19338 • Published Jun 17 • 52
WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation Paper • 2605.10912 • Published May 11 • 48
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Paper • 2605.20110 • Published May 19 • 4
DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders Paper • 2605.22777 • Published May 21 • 5
LoMo: Local Modality Substitution for Deeper Vision-Language Fusion Paper • 2605.30265 • Published May 28 • 24
Not only where, But when: Temporal Scheduling for RLVR Paper • 2605.25381 • Published May 25 • 6