Skip to content
Make AI Good

Graph · Person

Neel Nanda

01 · In focus

One person, in the field.

The structured facts the source records about Neel Nanda, the count of declared adjacencies in the corpus, and the federation map zoomed on this node and its neighbours.

person

1 declared connection

Kind
Person
Status
active
Confidence
high
Entity ID
person-neel-nanda
Network
View in network

Tags united-kingdom, ai-safety, mechanistic-interpretability, interpretability, frontier-ai, google-deepmind, anthropic, right-to-warn, researcher

Neel Nanda · 1 direct neighbour visible

02 · Connections

1 adjacency, by relation.

Split by direction. Direct links are the ones Neel Nanda’s source record names; inferred backlinks are records elsewhere in the corpus that point at this entity.

Inferred backlinks

1 link

Other records that name this entity.

03 · Background

From the source record.

Body prose as it appears in movement-graph’s published markdown for this entity. Links to other corpus entities resolve to their graph page; links to deeper repo paths are kept as text so the page does not invent a route.

British AI safety researcher; Staff Research Scientist at Google DeepMind, where he leads the mechanistic interpretability team. He studied mathematics at the University of Cambridge (BA, 2020), after which he spent a gap year exploring AI safety through internships at the Future of Humanity Institute, DeepMind, and the Centre for Human-Compatible AI at UC Berkeley, before taking quantitative finance internship roles at Jane Street and Jump Trading.

He joined Anthropic as a language-model interpretability researcher under Chris Olah, then moved to Google DeepMind in 2022 to build and lead their mechanistic interpretability group. He created TransformerLens, a widely-used open-source library for studying the internal computations of language models, and runs a technical blog and YouTube channel on mechanistic interpretability research. In 2023 MIT Technology Review named him to its Innovators Under 35 list.

On 4 June 2024 he was one of seven researchers to sign "A Right to Warn about Advanced Artificial Intelligence" by name — alongside Jacob Hilton, Daniel Kokotajlo, Ramana Kumar, William Saunders, Carroll Wainwright, and Daniel Ziegler — and the only named current Google DeepMind employee among them. The letter, also endorsed by Geoffrey Hinton, Yoshua Bengio, and Stuart Russell, demanded structural whistleblower protections inside frontier-AI companies. The event is documented in the Right to Warn open letter event and the insider whistleblowing from frontier labs strategy entry.

04 · Sources

Where this came from.

5 sources listed from the pinned corpus. Links are shown only when the source URL is a valid HTTP(S) address.

  1. righttowarn.ai

    Checked 2026-06-12

    Original "A Right to Warn about Advanced Artificial Intelligence" letter (4 June 2024) — primary source for Nanda's status as a named co-signatory alongside six other researchers; also names DeepMind as his affiliation at time of signing

  2. time.com

    Checked 2026-06-12

    Time (4 June 2024) — "Employees Say OpenAI and Google DeepMind Are Hiding Dangers from the Public" — confirms Nanda's DeepMind affiliation and his status as the only named current DeepMind employee among the signatories

  3. neelnanda.io

    Checked 2026-06-12

    Personal website about page — primary source for Cambridge mathematics degree, career trajectory (FHI, DeepMind, CHAI internships; Anthropic; Google DeepMind), and current role as Staff Research Scientist and MI team lead

  4. innovatorsunder35.com

    Checked 2026-06-12

    MIT Technology Review Innovators Under 35 (2023) — secondary source for public recognition and Cambridge maths background

  5. 80000hours.org

    Checked 2026-06-12

    80,000 Hours podcast (part 2) — secondary source for career trajectory and role leading Google DeepMind's MI team at 26

Source: entities/persons/person-neel-nanda.md — movement-graph pin 5d136ad.