Pelly
Pellypp
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning upvoted a paper about 1 month ago
STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability upvoted a paper 3 months ago
Meta-CoT: Enhancing Granularity and Generalization in Image EditingOrganizations
None yet