The Singular Value Decompositions of Transformer Weight Matrices

mechanistic-interpretabilitysvdtransformersgpt2weight-analysis

Abstraction: SVD of GPT-2 weight matrices reveals interpretable semantic directions

Key points:

Connections: GPT-2 · Openai · Mechanistic Interpretability · Transformers · Superposition

Source: https://www.lesswrong.com/posts/mkbGjzxD8d8XqKHzA/the-singular-value-decompositions-of-transformer-weight