Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 6 days ago • 21
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 5 days ago • 142
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published 12 days ago • 141
Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading Paper • 2607.08964 • Published 19 days ago • 75
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training Paper • 2607.05804 • Published 21 days ago • 18
view article Article From Hugging Face to Amazon SageMaker Studio in one click amazon • 20 days ago • 17
Nemotron Agentic & Tool-Use Collection Datasets for building models capable of function calling, multi-step agentic tasks, terminal use, and SWE workflows. • 11 items • Updated 11 days ago • 13
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning Paper • 2607.07508 • Published 20 days ago • 26