NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
self-evolution
4 papers
Absolute Zero: Reinforced Self-play Reasoning with Zero Data
SEEA-R1: Tree-Structured Reinforcement Fine-Tuning for Self-Evolving Embodied Agents
Sparta Alignment: Collectively Aligning Multiple Language Models through Combat
TTRL: Test-Time Reinforcement Learning