Skip to content

Latest commit

 

History

15 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 

Repository files navigation

An collection of colab notebooks for various mechinterp miniprojects.

  • Measuring Superposition: An overview of the Linear representation hypothesis and superposition principle and a demo the SAE features for GPT-2 small. Methods: SAE, TransformerLens Open In Colab Notebook

  • Interpreting SAE features: A replicating methods of OpenAI and EleutherAI to generate and evaluate GPT-2 SAE feature interpretations. Methods:Transformerlens. Open In Colab Notebook

  • Transformer Circuit Analysis: A sampler for reverse-engineered attention heads in GPT-2-small to explain Induction Heads, replicating Anthropic’s findings. Methods: Activation patching, attention visualization (TransformerLens). notebook

About

Colab notebooks of various mechinterp mini projects

Resources

Stars

0 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages