NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
multi-layer perceptrons
3 papers
Exploring the Translation Mechanism of Large Language Models
Minimum Width for Deep, Narrow MLP: A Diffeomorphism Approach
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers