Evan Hubinger
Alignment researcher who co-wrote the paper naming mesa-optimization
An alignment researcher and lead author of “Risks from Learned Optimization”, the 2019 paper that named mesa-optimization and inner alignment: the problem of a trained system quietly pursuing its own objective inside the one you gave it.
In the library
Nothing of theirs is filed in the hub yet. Browse the full library.
Recommended next
Hand-picked from the hub based on what Evan Hubinger covers.
Partnership on AI — resource library
Partnership on AI's open library of guidance, frameworks, and case studies on responsible AI: synthetic media, labor and the economy, AI safety, fairness, and inclusive AI development.
Why this: Also about AI safety
UC Berkeley Center for Human-Compatible AI (CHAI)
This university research center focuses on creating safe and beneficial artificial intelligence. You can read published research papers, follow recent news and blog updates, and explore opportunities to work with their team.
Why this: Also about AI safety
Yoshua Bengio
Deep learning pioneer and Turing Award winner, now focused on AI risk. His site holds papers, talks and written positions; his Google Scholar list has the full publication record, most-cited first.
Why this: Also about AI safety
Superintelligence: The Idea That Eats Smart People
A skeptical 2016 talk by Maciej Ceglowski (Idle Words) that walks through the arguments for a superintelligence-driven intelligence explosion and takes them apart. Ceglowski compares the superintelligence risk debate to the Manhattan Project question of whether the first nuclear test could ignite the atmosphere, and argues that the core premises rest on speculative leaps rather than settled science. The talk was given at Web Camp Zagreb and is published as a full text transcript.
Why this: Also about AI safety
