transformer-lens-interpretability
Provides guidance for mechanistic interpretability research using TransformerLens to inspect and manipulate transformer internals via HookPoints and activation caching. Use when reverse-engineering model algorithms, studying attention patterns, or performing activation patching experiments.
- Version
- 1.0.0
- License
- MIT
Pinned to revision df088027ff23, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/transformer-lens-interpretability/SKILL.md
- skills/transformer-lens-interpretability/references/README.md
- skills/transformer-lens-interpretability/references/api.md
- skills/transformer-lens-interpretability/references/tutorials.md
Every link opens the file at its source, pinned to the revision this page describes.