NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Lin Yan
3 papers
ByteDance Inc.
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
EvaLearn: Quantifying the Learning Capability and Efficiency of LLMs via Sequential Problem Solving
Exploring Data Scaling Trends and Effects in Reinforcement Learning from Human Feedback