Accordion-Thinking: Self-Regulated Step Summaries for Efficient and Readable LLM Reasoning (https://arxiv.org/abs/2602.03249)
Zhicheng YANG
yangzhch6
AI & ML interests
reasoning with LLMs
Organizations
None yet
Acccordion-Thinking
Accordion-Thinking: Self-Regulated Step Summaries for Efficient and Readable LLM Reasoning (https://arxiv.org/abs/2602.03249)
Mirror-Critique
Critique to Verify: Accurate and Honest Test-Time Scaling with RL-Trained Verifiers (https://arxiv.org/abs/2509.23152)
DARS
Dataset & Model of [Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration](https://arxiv.org/abs/2508.13755v1)