NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
policy alignment
3 papers
Direct Alignment with Heterogeneous Preferences
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
Strategyproof Reinforcement Learning from Human Feedback