NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Lifan Yuan
3 papers
University of Illinois Urbana-Champaign
Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
TTRL: Test-Time Reinforcement Learning
The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning