Jan Leike

persona · active · confidence: medium · reviewed 2026-06-16

Affiliation: Ex-Anthropic alignment researcher (departed 2026); ex-OpenAI superalignment co-lead (resigned May 2024); PhD in machine learning One-line position: Alignment is the central technical challenge of our time, frontier labs are systematically under-investing in it, and the people closest to the problem keep leaving because they can't fix it from inside.

What he's reacting against

Key claims

Theories aligned with

Where he overlaps / splits

Notable predictions

Track record

Empirical vs normative

Sources

Weak spots / open questions

Rich's take

converts-from: personas/jan-leike.md · schema v1 · AI & Society domain