Nearby in the stack

Equivariant Sparse Autoencoders: Mechanistic Interpretability of Neural Networks on Symmetric Data · arXivDesk