NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
alignment approaches
3 papers
Improving LLM General Preference Alignment via Optimistic Online Mirror Descent
Reducing the Probability of Undesirable Outputs in Language Models Using Probabilistic Inference
Token-Level Self-Play with Importance-Aware Guidance for Large Language Models