NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
capabilities
3 papers
Fast attention mechanisms: a tale of parallelism
Probabilistic Token Alignment for Large Language Model Fusion
Why Do Some Language Models Fake Alignment While Others Don't?