NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
diagonal linear networks
3 papers
Adam Reduces a Unique Form of Sharpness: Theoretical Insights Near the Minimizer Manifold
Alternating Gradient Flows: A Theory of Feature Learning in Two-layer Neural Networks
Heavy-Ball Momentum Method in Continuous Time and Discretization Error Analysis