NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Tom Griffiths
3 papers
Princeton University
Are Large Language Models Sensitive to the Motives Behind Communication?
Causal Head Gating: A Framework for Interpreting Roles of Attention Heads in Transformers
Partner Modelling Emerges in Recurrent Agents (But Only When It Matters)