NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
co-evolution
3 papers
CURE: Co-Evolving Coders and Unit Testers via Reinforcement Learning
Is PRM Necessary? Problem-Solving RL Implicitly Induces PRM Capability in LLMs
RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning