NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Csaba Szepesvari
3 papers
Google DeepMind / University of Alberta
Beyond Least Squares: Uniform Approximation and the Hidden Cost of Misspecification
Eluder dimension: localise it!
REINFORCE Converges to Optimal Policies with Any Learning Rate