NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
online rlhf
3 papers
Ask a Strong LLM Judge when Your Reward Model is Uncertain
Avoiding exp(R) scaling in RLHF through Preference-based Exploration
Provably Efficient Online RLHF with One-Pass Reward Modeling