YFarmX

Sparse autoencoder (SAE)

AI

Sparse autoencoder (SAE): An interpretability tool that decomposes a model's dense activations into many sparse features, each hopefully corresponding to a single human-understandable concept.

Related terms

Browse the full glossary →