NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
mechanistic analysis
3 papers
Bigram Subnetworks: Mapping to Next Tokens in Transformer Language Models
Born a Transformer -- Always a Transformer? On the Effect of Pretraining on Architectural Abilities
Too Late to Recall: Explaining the Two-Hop Problem in Multimodal Knowledge Retrieval