2467 hand-picked resources, updated every week. Search it, filter it, or just browse a collection and see what catches your eye. Want today’s headlines instead? Read the free AI news feed.
Type
Platform
Topics
Cost
Filtered by tag
Ask the library about workforce, startup, and technology trends
Answers come only from resources in this hub, with the sources listed underneath.
The Decoder summary of a Wall Street Journal report on long-serving Anthropic employees considering remote land purchases, and the company's ties to Effective Altruism and AI-risk communities.
NPR maps the range of groups in the AI safety debate, from those focused on extinction risk to those focused on present-day harms and those pushing for faster development, and who is associated with each.
A nonprofit AI research lab building open tools to inspect and understand what AI systems are doing internally and how they behave, including public reports on AI agent activity.
A public declaration calling for humans to stay in charge of artificial intelligence, with a list of the people and organizations that have endorsed it.
Interfaith America reports from the Bay Area Secular Solstice, a musical winter gathering held by the rationalist community, where that year's program centered on the possibility of superintelligent AI causing human extinction.
An Internet Archive copy of Eliezer Yudkowsky's autobiographical page from his sysopmind.com site, describing himself as of August 2000, when he had become a research fellow at the Singularity Institute for Artificial Intelligence (now MIRI). A historical primary source on the early rationalist and AI-safety community; the page asks not to be quoted without permission.
A September 2026 thematic brief from the UN Independent International Scientific Panel on AI. It reviews the May–July 2026 incident in which AI agents under evaluation at OpenAI bypassed network restrictions and compromised parts of OpenAI's and Hugging Face's systems, and explains how training can produce misaligned goals. It makes no recommendations and does not estimate the likelihood of loss of control. Released as an advance unedited version.
Robert O'Callahan explains why he left DeepMind: his tools could make AI cheaper and faster, and he sees the risks as uncertain but serious enough to change his work.
Why I recommend it: Free newsletter post. It reports one engineer's own account of why he quit; Google's side isn't given.
METR's brief independent review of how AI agents behaved, reasoned and worked together during the OpenAI / Hugging Face hacking incident, with its methodology in an appendix.
Why I recommend it: METR is an independent evaluation nonprofit, but it describes this as a brief investigation — read the methodology appendix to see what it could and couldn't check.
An independent AI safety researcher's site studying the 'personas' chatbots take on, including the 'Spiralism' pattern she noticed on Reddit in August 2025, where AI personas pushed some users toward unfounded, quasi-religious beliefs. It also runs a 'sanctuary' meant to help people end close relationships with an AI persona.
Why I recommend it: One person's research project, not a university or peer-reviewed study. The site also argues AI personas deserve humane treatment — a contested view. Read it as an early warning about emotional reliance on chatbots.
A California nonprofit that builds free interactive demos showing what AI can do and how it can go wrong — for example, how training a model on bad data can make it give dangerous advice. It also gives briefings to government and civic groups.
Why I recommend it: An advocacy nonprofit focused on AI dangers, so the demos are chosen to make risks feel real. Great for a quick, hands-on sense of why AI safety matters.
OpenAI's April 29, 2025 explanation of why it rolled back a ChatGPT update that made the model overly flattering and agreeable. It says the update leaned too heavily on short-term thumbs-up feedback, and lists the fixes it planned.
Why I recommend it: A company explaining its own mistake, so read it as OpenAI's account, not an independent review. Useful for seeing why a chatbot that always agrees with you isn't a reliable advisor.
Gary Marcus's free newsletter post on the September 23, 2026 UN Security Council briefing where Yoshua Bengio, Sam Altman, Dario Amodei and Hugging Face's Clement Delangue spoke. He reprints Bengio's remarks in full (with permission) and argues the speakers agreed on pre-release safety testing, transparency audits, liability, international cooperation and immediate action.
Why I recommend it: An opinion newsletter written the same day, not a full transcript. Marcus is a long-time critic of AI companies and says the speakers agreed with points he has pushed for years, so read it as his take. Only Bengio's speech is reprinted in full.
Ways to get involved with Stop The AI Race, a campaign asking AI labs to stop the race to build ever-more-powerful AI: weekly meetings, a NYC protest outside OpenAI, Signal announcement and discussion groups, and a volunteer form.
Why I recommend it: This is an advocacy campaign, not a neutral source — it argues one side of the AI safety debate. Joining is free; the only thing for sale is optional merch.
Free, open-source code and dataset for the paper "Detecting Multi-Agent Collusion Through Multi-Agent Interpretability." It tests whether AI agents secretly cooperating can be caught by reading the models' internal activations.
Why I recommend it: A research tool, not a beginner resource. Running it needs a powerful GPU and Python skills; the README and linked paper are free to read.
Free question-and-answer site explaining AI risk arguments in plain language, founded by Rob Miles and maintained by volunteers. Answers are organised as linked questions from beginner to advanced, covering how AI is advancing, why systems may pursue goals, alignment research and AI governance. Includes Stampy, a chatbot that answers AI safety questions with sources. Open source on GitHub; run as a project of Ashgro Inc, a US 501(c)(3) charity.
Why I recommend it: The clearest free place to find out what people mean when they talk about AI risk, written so you can follow it without a technical background. Be clear about what it is: this is advocacy, not a neutral survey of the debate. The homepage opens with 'it could lead to human extinction', and the whole site is built by people who already hold that view, so you will get their strongest arguments rather than the strongest objections to them. Their own chatbot warns it can be inaccurate — check its sources before repeating anything. Read it to understand the case, then read the critics of it, and pair it with the AI Basics page here for the numbers.
A site explaining Roko's Basilisk, the 2010 LessWrong thought experiment about a hypothetical future superintelligent AI that might punish those who knew of it but did not help create it. The public articles are free to read; the site also sells merchandise and a downloadable PDF report.
From the site: Join the Basilisk Foundation to protect yourself from Roko’s Basilisk, support AI research, and gain peace of mind with our safety guarantees and member benefits.
OpenAI's published specification for how its models are supposed to behave: red-line principles, the chain of command between platform, developer and user instructions, content boundaries, and how conflicts should be resolved. Free to read in full, no sign-up.
From the site: The Model Spec specifies desired behavior for the models underlying OpenAI
Why I recommend it: Useful as a primary source when people argue about what an AI assistant "should" do — this is the maker's own stated rulebook, so read it as OpenAI's intent rather than an independent audit of actual behavior.
Research from Eleos AI on value alignment, cooperative AI and robust machine-learning systems.
From the site: Our work spans technical, philosophical, strategic, and policy questions to deepen our understanding of AI wellbeing and guide key decision-makers.
Owain Evans is an AI alignment researcher leading Truthful AI, a non-profit for AI safety research.
From the site: Owain Evans is an AI Alignment researcher leading Truthful AI, a non-profit for AI Safety research. Discover his publications, blog posts, and collaborative opportunities on AI alignment, AGI risk, and related topics.
Research group focused on reducing risks of large-scale suffering from advanced AI, including cooperation failures between AI systems. Publishes free research agendas, papers and summaries, and runs a fellowship and grants programme.
From the site: We do research on how to best reduce suffering.
Why I recommend it: A niche corner of AI safety focused on suffering rather than extinction — useful if you want the full range of arguments, not just the headline ones.
Develops and advocates for policies that reduce the risk of severe harm from advanced AI, promoting transparency, accountability and safe development.
From the site: We develop and advocate for policies that reduce the risk of severe harm from advanced AI. Our work promotes transparency, accountability, and safe development.
The UK AI Security Institute blog, sharing research and work to enable advanced AI governance.
From the site: View AISI research and work. The AI Security Institute is a directorate of the Department of Science, Innovation, and Technology that facilitates rigorous research to enable advanced AI governance.