FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience Paper • 2609.03241 • Published 7 days ago • 87
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations Paper • 2608.15930 • Published 25 days ago • 47
Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning Paper • 2607.18722 • Published Jul 21 • 36
Agentic Troubleshooting Guide Automation for Incident Management Paper • 2510.10074 • Published Oct 11, 2025 • 3
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published Jul 14 • 235
Group-in-Group Policy Optimization for LLM Agent Training Paper • 2505.10978 • Published May 16, 2025 • 24
AEM: Adaptive Entropy Modulation for Multi-Turn Agentic Reinforcement Learning Paper • 2605.00425 • Published May 8 • 23
Qianfan-VL: Domain-Enhanced Universal Vision-Language Models Paper • 2509.18189 • Published Sep 19, 2025 • 4
Agentic Troubleshooting Guide Automation for Incident Management Paper • 2510.10074 • Published Oct 11, 2025 • 3
A Survey of Context Engineering for Large Language Models Paper • 2507.13334 • Published Jul 17, 2025 • 263