TextReg: Mitigating Prompt Distributional Overfitting via Regularized Text-Space Optimization Paper • 2605.21318 • Published 3 days ago • 8
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 21 days ago • 139
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published about 1 month ago • 376
Steering Geometry: Validating Human Value Geometry in LLM Steering Space Paper • 2609.06289 • Published Sep 5 • 31
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published Sep 3 • 28
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published Sep 3 • 186
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published Sep 2 • 408
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 266
Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval Paper • 2608.06060 • Published Aug 6 • 41
UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Paper • 2608.04701 • Published Aug 5 • 9
StyleForge: Indoor Furniture Styling by Counterfactual Reasoning in a Hypergraph Field Paper • 2608.01954 • Published Aug 3 • 13