
Eliezer Yudkowsky
AI safety researcher; co-founder, MIRI
Among the earliest and most prominent voices warning about risks from advanced AI. Writes at length on LessWrong, the forum he helped start.
Key arguments & positions
- Argues building superintelligence without solving alignment first will very likely be catastrophic, and does not believe current safety techniques scale.
- Advocates an international, verified moratorium on frontier AI training runs.
- Popularized foundational AI-safety concepts such as "coherent extrapolated volition."
Accomplishments
- Co-founded the Machine Intelligence Research Institute (MIRI) in 2000.
- Founded LessWrong in 2009.
- Co-authored the 2025 New York Times bestseller "If Anyone Builds It, Everyone Dies."
Papers & key writings
- "If Anyone Builds It, Everyone Dies" (2025, with Nate Soares)
- "Coherent Extrapolated Volition" (2004, MIRI)
Links
Timeline
- 2000Co-founds the institute that becomes MIRI.
- 2009Begins the Sequences on rationality, later collected as a free book.
- 2010Publishes Harry Potter and the Methods of Rationality.
- 2022Publishes "AGI Ruin: A List of Lethalities".
- 2023Argues in TIME for an international halt to frontier training runs.
In the library
Nothing of theirs is filed in the hub yet. Browse the full library.
Recommended next
Hand-picked from the hub based on what Eliezer Yudkowsky covers.
METR
A research nonprofit that independently evaluates frontier AI models to measure what they can actually do and what risks that creates. Reports are free.
Why this: Covers AI safety and ethics too
These Are the Most Urgent AI Risks, According to 272 Experts
MIT Sloan summary of research surveying 272 experts on which AI risks could cause the most harm in the next five years.
Why this: Covers AI safety and AI risk too
PyRIT
Microsoft's open-source toolkit for red-teaming AI systems: automated attack prompts, scoring of the responses, and repeatable runs. Free.
Why this: Covers AI safety and AI risk too
Concrete Problems in AI Safety
The 2016 paper that framed AI safety as a set of specific engineering problems — side effects, reward hacking, unsafe exploration — rather than a philosophical worry. Free on arXiv.
Why this: Covers AI safety and AI risk too
