GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation Paper • 2609.05588 • Published 10 days ago • 58
MNIST-PRO: MNIST is Back as a Partially Observable World for AI Agents Paper • 2608.31022 • Published 14 days ago • 9
ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step Paper • 2608.02358 • Published Aug 3 • 11
Data Agent: Learning to Select Data via End-to-End Dynamic Optimization Paper • 2603.07433 • Published May 13
$Σ$-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems Paper • 2607.27958 • Published Jul 30 • 17
IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation Paper • 2607.22375 • Published Jul 24 • 9
On the Limits of LLM-as-Judge for Scientific Novelty Assessment Paper • 2606.12071 • Published Jun 10 • 3
GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards Paper • 2606.04889 • Published Jun 3 • 4
Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Paper • 2605.06165 • Published May 7 • 1
$δ$-mem: Efficient Online Memory for Large Language Models Paper • 2605.12357 • Published May 12 • 133
DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Paper • 2505.17795 • Published May 23, 2025
Stacked from One: Multi-Scale Self-Injection for Context Window Extension Paper • 2603.04759 • Published Apr 9
Epistemic Context Learning: Building Trust the Right Way in LLM-Based Multi-Agent Systems Paper • 2601.21742 • Published Jan 29
From Perception to Action: An Interactive Benchmark for Vision Reasoning Paper • 2602.21015 • Published Feb 24 • 26
Bi-Bimodal Modality Fusion for Correlation-Controlled Multimodal Sentiment Analysis Paper • 2107.13669 • Published Jul 28, 2021
Error-Free Linear Attention is a Free Lunch: Exact Solution from Continuous-Time Dynamics Paper • 2512.12602 • Published Dec 14, 2025 • 44
NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards Paper • 2511.14659 • Published Nov 18, 2025 • 13
10 Open Challenges Steering the Future of Vision-Language-Action Models Paper • 2511.05936 • Published Nov 8, 2025 • 6
Demystifying deep search: a holistic evaluation with hint-free multi-hop questions and factorised metrics Paper • 2510.05137 • Published Oct 1, 2025 • 6
OffTopicEval: When Large Language Models Enter the Wrong Chat, Almost Always! Paper • 2509.26495 • Published Sep 30, 2025 • 13