arxiv:2607.23802
Kun Wan
timecuriosity
AI & ML interests
None yet
Recent Activity
authored a paper about 13 hours ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement upvoted a paper about 15 hours ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-ImprovementOrganizations
None yet