News

Jul 2026 Ana Lucic has been awarded an NWO Veni grant to work on mechanistic interpretability for graph ML models!
Jun 2026 We have one paper accepted to the Mechanistic Interpretability workshop at ICML 2026 on investigating what routers learn in Mixture-of-Depths models.
Feb 2026 We have one paper on investigating atmospheric structure in Aurora accepted to the Scientific Methods for Understanding Deep Learning workshop at ICLR 2026!
Nov 2025 We have two papers accepted to the Mechanistic Interpretability workshop at NeurIPS 2025! One is about equivariant sparse autoencoders, the other is about causal abstraction as a framework for faithfulness.
May 2025 Ege Erdogan has joined us as a new PhD student working on mechanistic interpretability. Welcome Ege!
Nov 2024 We’re hiring! We have a PhD position available at the UvA on mechanistic interpretability.
Nov 2024 Ana Lucic has returned to the University of Amsterdam as an assistant professor! Looking forward to seeing old faces and meeting new ones.