NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Jiaxuan Gao
3 papers
IIIS, Tsinghua University
AREAL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning
How Far Are We from Optimal Reasoning Efficiency?
Reasoning Is Not a Race: When Stopping Early Beats Going Deeper