Artificial Atlas

Who is betting on what in AI, and against whom.

CO

Chris Olah

Researcher

Anthropic co-founder and the leading figure in mechanistic interpretability, the effort to read what is happening inside a trained network. Interpretability is the safety layer inside camp 1.

Edit this entry

Works in

  1. Jan 2021 – now
    Anthropic, Co-founder; interpretability lead
    Sources Chris Olah Wikipedia 2026-09-04