arxiv:2606.04923
Zhuoyuan Hao 郝卓远
larry2210
AI & ML interests
LLM reasoning, reinforcement learning, reward hacking, trustworthy AI, AI agents
Recent Activity
authored a paper about 2 months ago
Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning upvoted a paper about 2 months ago
Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning submitted a paper about 2 months ago
Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement LearningOrganizations
None yet