Last released Jan 10, 2026
A library to train and use Sparse Autoencoders for mechanistic interpretability
Supported by