Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training Paper • 2609.07108 • Published 4 days ago • 27
uno Collection Unlocking Lossless Speedups in LLMs via Discrete Diffusion • 4 items • Updated 3 days ago • 4
view article Article NeoMME: an efficient Multimodal-native and Multilingual Encoder Hcompany • 7 days ago • 90
view article Article Internet-Scale Knowledge Retrieval: A Novel Vector Search Dataset at 10B Scale Qdrant • 9 days ago • 14
Training Agents to Evolve with Their Harness: TaoLive Digital Avatar Agent Technical Report Paper • 2608.15763 • Published 20 days ago • 54
GenFirst: Generation Before Reconstruction for Stable End-to-End Latent Generative Modeling Paper • 2608.29335 • Published 13 days ago • 69
Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO Paper • 2608.27351 • Published 15 days ago • 22
Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning Paper • 2608.23318 • Published 18 days ago • 32
GigaBrain-0.7: Scaling Embodied Foundation Models to Emergent Capabilities with a Three-System Architecture Paper • 2608.15875 • Published 26 days ago • 103
Apodex 1.1: Scaling Agentic Intelligence for Complex Work Paper • 2608.23283 • Published 18 days ago • 206
ANEForge: Python for direct computation on the Apple Neural Engine Paper • 2606.17090 • Published Jun 12 • 3
Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL Paper • 2608.17253 • Published 23 days ago • 97
Second Thought: Reasoning in Parallel as LLM Agents Act and Observe Paper • 2608.13667 • Published 28 days ago • 17
Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution Paper • 2608.08311 • Published Aug 8 • 91
Next-Latent Prediction Transformers Learn Compact World Models Paper • 2511.05963 • Published Nov 8, 2025 • 5
OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents Paper • 2608.05013 • Published Aug 4 • 37
view article Article Training a coding agent using the OpenCode harness in remote HF sandboxes with TRL and OpenEnv sergiopaniego • Aug 5 • 26