arxiv:2606.14885
ZhuofengLi
ZhuofengLi
AI & ML interests
Agents, Reasoning LLMs/VLLMs, RL
Recent Activity
upvoted a paper about 6 hours ago
EasyPPO: Stabilizing the Critic Is Key updated a dataset 16 days ago
ZhuofengLi/Harbor-SWE-Trajectory published a dataset 17 days ago
ZhuofengLi/Harbor-SWE-Trajectory