DepthBench: Measuring How Residual Connections Enable More Computational Depth Paper • 2609.32534 • Published 9 days ago • 32
Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR Paper • 2609.37868 • Published 6 days ago • 60
Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation Paper • 2609.35347 • Published 7 days ago • 178
Do Implicit Personalization and Explicit Styles Conflict? PsPLUG: A Lightweight Plug-in for Balancing Personalization and Style in Customized LLMs Paper • 2601.06362 • Published 15 days ago • 10
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 10 days ago • 159
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 12 days ago • 43