sparse-autoencoder-training
Provides guidance for training and analyzing Sparse Autoencoders (SAEs) using SAELens to decompose neural network activations into interpretable features. Use when discovering interpretable features, analyzing superposition, or studying monosemantic representations in language models.
- Version
- 1.0.0
- License
- MIT
Pinned to revision df088027ff23, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/sparse-autoencoder-training/SKILL.md
- skills/sparse-autoencoder-training/references/README.md
- skills/sparse-autoencoder-training/references/api.md
- skills/sparse-autoencoder-training/references/tutorials.md
Every link opens the file at its source, pinned to the revision this page describes.