Chris Olah; Arvind Satyanarayan; Ian Johnson; Shan Carter; Ludwig Schubert; Katherine Ye; Alexander Mordvintsev · Google Brain
The founding text for interactive interpretability. Argues that interpretability techniques studied in isolation are far weaker than the interfaces you get by composing them, and demonstrates this with a set of live, hoverable interfaces over an image classifier.
Established that the interface IS the contribution — the template every entry in this ledger inherits from.
High · Read in full (article + Distill metadata), 2026-08-06
open full entry