Sparse Autoencoders

Recent momentum

-60%

8 papers in the last 28 days · 0.2% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-14

7 new papers

A weekly snapshot of new work published in Sparse Autoencoders.

153 papers

Latest in Sparse Autoencoders

Open your feed →
CardsList
  1. LLM Layers Immediately Correct Each Other

    Sep 7, 2026Arjun Patrawala, Jiahai Feng, Erik Jones +1Model ActivationsResidual Stream

  2. Probing and steering biology across Boltz-1s trunk-diffusion boundary

    Aug 11, 2026Piotr Jedryszek, Tongmeng Xie, Adam Winnifrith +5ProteinFeature Alignment

  3. "Many Are My Names": The Anatomy of the Assistant and Its Personas via Sparse Autoencoders

    Aug 8, 2026Adelaide Danilov, Aria Nourbakhsh, Oleksandr Marchenko Breneur +1PersonasRole-Playing

  4. PairSAE: Mechanistic Interpretability from Pair Representations in Protein Co-Folding

    Jun 25, 2026Giosue Migliorini, Aristofanis Rontogiannis, Grigori Guitchounts +3ProteinSparse Autoencoders

  5. Do Sparse Autoencoders Learn Meaningful Concept Hierarchies?

    Jun 22, 2026Nils Grandien, David Steinmann, Felix Friedrich +1Sparse AutoencodersHierarchical