NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
best-of-n sampling
3 papers
Does Thinking More Always Help? Mirage of Test-Time Scaling in Reasoning Models
LASeR: Learning to Adaptively Select Reward Models with Multi-Arm Bandits
Rethinking Optimal Verification Granularity for Compute-Efficient Test-Time Scaling