arxiv:2512.13095
xinhong ma
stefanxinhong
ยท
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
Learning from Your Own Mistakes: Constructing Learnable Micro-Reflective Trajectories for Self-Distillation authored a paper about 2 months ago
ADHint: Adaptive Hints with Difficulty Priors for Reinforcement Learning upvoted a paper about 2 months ago
ADHint: Adaptive Hints with Difficulty Priors for Reinforcement LearningOrganizations
None yet