INTERP ATLAS

year

2019

1 entry

2019 (1)

IA-007

Activation Atlas is and showing , about .

Activation Atlas
2026-08-12

Shan Carter; Zan Armstrong; Ludwig Schubert; Ian Johnson; Chris Olah · Google Brain; OpenAI

Renders millions of activations from an image classifier as feature-inversion images laid out on a single navigable map, so you can pan across the concepts a network has learned the way you would read an atlas.

Made a model's whole learned concept space visible at once rather than one neuron at a time — the clearest ancestor of the feature-neighbourhood maps in Scaling Monosemanticity.

HighRead the Distill article and its citation metadata, 2026-08-06

open full entry