arxiv:2601.05167
Langlin Huang
shrango
AI & ML interests
LLM Reasoning, Machine Translation
Recent Activity
upvoted a paper 1 day ago
Hermes: Learning Contextual Reasoning Unlocks Test-Time Scaling upvoted a paper 1 day ago
MILO: Automated Harness Discovery via Orchestrated Multi-Agent Evolution upvoted a paper 1 day ago
On the Off-Policy Teacher in On-Policy Distillation