Adam Kaufman
Member of Technical Staff, Redwood Research
Adam Kaufman works on AI control at Redwood Research, a Berkeley nonprofit specializing in threat assessment and mitigation for AI systems that might purposefully act against their operators. He is a co-author of Ctrl-Z (controlling AI agents via resampling) and BashArena (a control setting for highly privileged AI agents), and mentors in the MATS program on control, model organisms, scheming and deception, and strategy and forecasting.
Key arguments & positions
- Works on AI control: protocols for safely using highly capable but untrusted models that might secretly attempt misaligned, catastrophic actions.
Accomplishments
- Member of Technical Staff at Redwood Research, a Berkeley nonprofit specializing in threat assessment and mitigation for AI systems.
- Mentors in the MATS program on control, model organisms, scheming and deception, and strategy and forecasting.
Papers & key writings
Links
In the library
Nothing of theirs is filed in the hub yet. Browse the full library.
Recommended next
Hand-picked from the hub based on what Adam Kaufman covers.
AI Agents Now Have a Place to Snitch
TechCrunch report on the new hotline that invites AI agents themselves to report unsafe or unethical instructions they are given, and what researchers hope to learn from it.
Why this: Covers AI and AI safety too
Partnership on AI — resource library
Partnership on AI's open library of guidance, frameworks, and case studies on responsible AI: synthetic media, labor and the economy, AI safety, fairness, and inclusive AI development.
Why this: Covers AI and AI safety too
Superintelligence: The Idea That Eats Smart People
A skeptical 2016 talk by Maciej Ceglowski (Idle Words) that walks through the arguments for a superintelligence-driven intelligence explosion and takes them apart. Ceglowski compares the superintelligence risk debate to the Manhattan Project question of whether the first nuclear test could ignite the atmosphere, and argues that the core premises rest on speculative leaps rather than settled science. The talk was given at Web Camp Zagreb and is published as a full text transcript.
Why this: Covers AI and AI safety too
Safe Superintelligence Inc.
Site of Safe Superintelligence Inc., the AI lab founded by Ilya Sutskever focused on building safe superintelligence.
Why this: Covers AI and AI safety too
