ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments Paper • 2609.19134 • Published 1 day ago • 64
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 7 days ago • 34
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 9 days ago • 419