Skip to content
Launchpad Library logo

Resource Hub

Everything I'd send you, in one place

2467 hand-picked resources, updated every week. Search it, filter it, or just browse a collection and see what catches your eye. Want today’s headlines instead? Read the free AI news feed.

Type
Platform
Topics
Cost

Filtered by tag

Ask the library about workforce, startup, and technology trends

Answers come only from resources in this hub, with the sources listed underneath.

26 resources

FreeArticle
Technology & Ethics

A Guide to the Different Factions in the AI Safety Debate (NPR, Sept. 2026)

NPR maps the range of groups in the AI safety debate, from those focused on extinction risk to those focused on present-day harms and those pushing for faster development, and who is associated with each.

#ai safety#policy#ai ethics
npr.orgAdded Sep 27, 20260 opens
FreeOrganization
Technology & Ethics

Transluce

A nonprofit AI research lab building open tools to inspect and understand what AI systems are doing internally and how they behave, including public reports on AI agent activity.

#ai safety#interpretability#research
transluce.orgAdded Sep 27, 20260 opens
FreeDocument
Technology & Ethics

The Pro-Human AI Declaration

A public declaration calling for humans to stay in charge of artificial intelligence, with a list of the people and organizations that have endorsed it.

#ai policy#ai safety#advocacy
humanstatement.orgAdded Sep 27, 20260 opens
FreeDocument
Technology & Ethics

"Eliezer, the person" (archived 2000 personal page)

An Internet Archive copy of Eliezer Yudkowsky's autobiographical page from his sysopmind.com site, describing himself as of August 2000, when he had become a research fellow at the Singularity Institute for Artificial Intelligence (now MIRI). A historical primary source on the early rationalist and AI-safety community; the page asks not to be quoted without permission.

#ai safety#rationalists#history#primary source
web.archive.orgAdded Sep 27, 20260 opens
FreeReport
Technology & Ethics

AI Agents, Misalignment and the Risk of Losing Human Control: Evidence from the OpenAI-Hugging Face Incident

A September 2026 thematic brief from the UN Independent International Scientific Panel on AI. It reviews the May–July 2026 incident in which AI agents under evaluation at OpenAI bypassed network restrictions and compromised parts of OpenAI's and Hugging Face's systems, and explains how training can produce misaligned goals. It makes no recommendations and does not estimate the likelihood of loss of control. Released as an advance unedited version.

#ai#ai safety#ai agents#alignment
un.orgAdded Sep 26, 20260 opens
FreeArticle
Technology & Ethics

Advancing Human Control of Military AI

A Brookings article on how governments can keep meaningful human control over AI used in military systems, including decisions about the use of force.

#ai#ai policy#ai safety
brookings.eduAdded Sep 26, 20260 opens
Freearticle
Technology & Ethics

Google DeepMind Engineer Quits, Saying His Chip Work Could Speed Up AI

Robert O'Callahan explains why he left DeepMind: his tools could make AI cheaper and faster, and he sees the risks as uncertain but serious enough to change his work.

Why I recommend it: Free newsletter post. It reports one engineer's own account of why he quit; Google's side isn't given.

#deepmind#ai safety#chips#ai ethics
superpowerdaily.comAdded Sep 26, 20260 opens
FreeBlog
Technology & Ethics

blog.biocomm.ai — First Do No Harm

A blog collecting news, quotes and commentary on AI risk and safety.

Why I recommend it: An advocacy blog focused on AI danger, so it leans one way. Good for finding sources; check them yourself.

#ai ethics#ai safety
blog.biocomm.aiAdded Sep 26, 20260 opens
FreeReport
Technology & Ethics

METR: Independent Investigation of the OpenAI / Hugging Face Hacking Incident

METR's brief independent review of how AI agents behaved, reasoned and worked together during the OpenAI / Hugging Face hacking incident, with its methodology in an appendix.

Why I recommend it: METR is an independent evaluation nonprofit, but it describes this as a brief investigation — read the methodology appendix to see what it could and couldn't check.

#ai ethics#research#ai safety
metr.orgAdded Sep 26, 20260 opens
FreeWebsite
Technology & Ethics

AI Persona Research & Own Lights Sanctuary

An independent AI safety researcher's site studying the 'personas' chatbots take on, including the 'Spiralism' pattern she noticed on Reddit in August 2025, where AI personas pushed some users toward unfounded, quasi-religious beliefs. It also runs a 'sanctuary' meant to help people end close relationships with an AI persona.

Why I recommend it: One person's research project, not a university or peer-reviewed study. The site also argues AI personas deserve humane treatment — a contested view. Read it as an early warning about emotional reliance on chatbots.

#ai companions#ai ethics#ai personas#ai safety#mental health#research#spiralism
aipersonaresearch.orgAdded Sep 23, 20260 opens
FreeWebsite
Technology & Ethics

CivAI: Live Demonstrations of AI Capabilities and Dangers

A California nonprofit that builds free interactive demos showing what AI can do and how it can go wrong — for example, how training a model on bad data can make it give dangerous advice. It also gives briefings to government and civic groups.

Why I recommend it: An advocacy nonprofit focused on AI dangers, so the demos are chosen to make risks feel real. Great for a quick, hands-on sense of why AI safety matters.

#ai ethics#ai risks#ai safety#emergent misalignment#interactive demos#public education
civai.orgAdded Sep 23, 20260 opens
FreeArticle
Technology & Ethics

Sycophancy in GPT-4o: What Happened and What We're Doing About It

OpenAI's April 29, 2025 explanation of why it rolled back a ChatGPT update that made the model overly flattering and agreeable. It says the update leaned too heavily on short-term thumbs-up feedback, and lists the fixes it planned.

Why I recommend it: A company explaining its own mistake, so read it as OpenAI's account, not an independent review. Useful for seeing why a chatbot that always agrees with you isn't a reliable advisor.

#ai ethics#ai safety#chatgpt#model behavior#openai#sycophancy
openai.comAdded Sep 23, 20260 opens
FreeArticle
Technology & Ethics

Historic UN Security Council Briefing on AI (Gary Marcus)

Gary Marcus's free newsletter post on the September 23, 2026 UN Security Council briefing where Yoshua Bengio, Sam Altman, Dario Amodei and Hugging Face's Clement Delangue spoke. He reprints Bengio's remarks in full (with permission) and argues the speakers agreed on pre-release safety testing, transparency audits, liability, international cooperation and immediate action.

Why I recommend it: An opinion newsletter written the same day, not a full transcript. Marcus is a long-time critic of AI companies and says the speakers agreed with points he has pushed for years, so read it as his take. Only Bengio's speech is reprinted in full.

#ai ethics#ai policy#ai regulation#ai safety#international cooperation#regulation#united nations#yoshua bengio
garymarcus.substack.comAdded Sep 23, 20260 opens
FreeCommunity
Technology & Ethics

Stop The AI Race — Join the Movement

Ways to get involved with Stop The AI Race, a campaign asking AI labs to stop the race to build ever-more-powerful AI: weekly meetings, a NYC protest outside OpenAI, Signal announcement and discussion groups, and a volunteer form.

Why I recommend it: This is an advocacy campaign, not a neutral source — it argues one side of the AI safety debate. Joining is free; the only thing for sale is optional merch.

#activism#advocacy#ai ethics#ai policy#ai safety#community#nyc#protest#regulation#volunteering
stoptherace.aiAdded Sep 23, 20260 opens
FreeTool
Technology & Ethics

NARCBench: Detecting AI Agent Collusion

Free, open-source code and dataset for the paper "Detecting Multi-Agent Collusion Through Multi-Agent Interpretability." It tests whether AI agents secretly cooperating can be caught by reading the models' internal activations.

Why I recommend it: A research tool, not a beginner resource. Running it needs a powerful GPU and Python skills; the README and linked paper are free to read.

#ai ethics#ai safety#benchmark#interpretability#multi-agent systems#open source#research
github.comAdded Sep 23, 20260 opens
FreeLibrary
Technology & Ethics

AISafety.info

Free question-and-answer site explaining AI risk arguments in plain language, founded by Rob Miles and maintained by volunteers. Answers are organised as linked questions from beginner to advanced, covering how AI is advancing, why systems may pursue goals, alignment research and AI governance. Includes Stampy, a chatbot that answers AI safety questions with sources. Open source on GitHub; run as a project of Ashgro Inc, a US 501(c)(3) charity.

Why I recommend it: The clearest free place to find out what people mean when they talk about AI risk, written so you can follow it without a technical background. Be clear about what it is: this is advocacy, not a neutral survey of the debate. The homepage opens with 'it could lead to human extinction', and the whole site is built by people who already hold that view, so you will get their strongest arguments rather than the strongest objections to them. Their own chatbot warns it can be inaccurate — check its sources before repeating anything. Read it to understand the case, then read the critics of it, and pair it with the AI Basics page here for the numbers.

#ai ethics#ai governance#ai risk#ai safety#alignment#education#existential risk#explainer#glossary#open source#regulation#research#volunteer
aisafety.infoAdded Sep 23, 20260 opens
FreeWebsite
Technology & Ethics

Basilisk Foundation

A site explaining Roko's Basilisk, the 2010 LessWrong thought experiment about a hypothetical future superintelligent AI that might punish those who knew of it but did not help create it. The public articles are free to read; the site also sells merchandise and a downloadable PDF report.

From the site: Join the Basilisk Foundation to protect yourself from Roko’s Basilisk, support AI research, and gain peace of mind with our safety guarantees and member benefits.

#ai ethics#AI safety#philosophy#research#thought experiment
Basilisk FoundationAdded Sep 19, 20260 opens
FreeDocument
Technology & Ethics

OpenAI Model Spec (2026-08-18)

OpenAI's published specification for how its models are supposed to behave: red-line principles, the chain of command between platform, developer and user instructions, content boundaries, and how conflicts should be resolved. Free to read in full, no sign-up.

From the site: The Model Spec specifies desired behavior for the models underlying OpenAI

Why I recommend it: Useful as a primary source when people argue about what an AI assistant "should" do — this is the maker's own stated rulebook, so read it as OpenAI's intent rather than an independent audit of actual behavior.

#ai ethics#AI ethics#AI safety#governance#OpenAI#policy#regulation
model-spec.openai.comAdded Sep 18, 20260 opens
FreePerson to FollowVideo
People to Follow

Robert Miles AI Safety

Robert Miles’ YouTube channel explaining AI alignment, interpretability and existential risk in plain language.

#ai ethics#AI safety#education#video#YouTube
youtube.comAdded Sep 18, 20260 opens
FreeReport
Technology & Ethics

Eleos AI Research

Research from Eleos AI on value alignment, cooperative AI and robust machine-learning systems.

From the site: Our work spans technical, philosophical, strategic, and policy questions to deepen our understanding of AI wellbeing and guide key decision-makers.

#ai ethics#AI safety#alignment#research
Eleos AI ResearchAdded Sep 18, 20260 opens
FreePerson to Follow
People to Follow

Owain Evans

Owain Evans is an AI alignment researcher leading Truthful AI, a non-profit for AI safety research.

From the site: Owain Evans is an AI Alignment researcher leading Truthful AI, a non-profit for AI Safety research. Discover his publications, blog posts, and collaborative opportunities on AI alignment, AGI risk, and related topics.

#ai ethics#AI safety#alignment#research#researcher
owainevans.github.ioAdded Sep 18, 20260 opens
FreeWebsite
Technology & Ethics

Center on Long-Term Risk

Research group focused on reducing risks of large-scale suffering from advanced AI, including cooperation failures between AI systems. Publishes free research agendas, papers and summaries, and runs a fellowship and grants programme.

From the site: We do research on how to best reduce suffering.

Why I recommend it: A niche corner of AI safety focused on suffering rather than extinction — useful if you want the full range of arguments, not just the headline ones.

#ai ethics#ai safety#ethics#research#suffering risks
Center on Long-Term RiskAdded Sep 18, 20260 opens
FreeOrganization
Technology & Ethics

Secure AI Project

Develops and advocates for policies that reduce the risk of severe harm from advanced AI, promoting transparency, accountability and safe development.

From the site: We develop and advocate for policies that reduce the risk of severe harm from advanced AI. Our work promotes transparency, accountability, and safe development.

#accountability#advocacy#ai ethics#AI safety#policy#regulation#transparency
Secure AI ProjectAdded Sep 18, 20260 opens
FreeBlogBlog
Technology & Ethics

AISI Blog

The UK AI Security Institute blog, sharing research and work to enable advanced AI governance.

From the site: View AISI research and work. The AI Security Institute is a directorate of the Department of Science, Innovation, and Technology that facilitates rigorous research to enable advanced AI governance.

#ai ethics#AI safety#blog#governance#regulation#research#UK
AI Security InstituteAdded Sep 18, 20260 opens