ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Paper • 2605.20278 • Published May 24 • 2
MaxProof: Scaling Mathematical Proof with Generative-Verifier RL and Population-Level Test-Time Scaling Paper • 2606.13473 • Published Jun 11 • 94
On the Limits of LLM-as-Judge for Scientific Novelty Assessment Paper • 2606.12071 • Published Jun 10 • 3