Skip to content
Launchpad Library logo

Resource Hub

Everything I'd send you, in one place

2893 hand-picked resources, updated every week. Search it, filter it, or just browse a collection and see what catches your eye. Want today’s headlines instead? Read the free AI news feed.

Type
Platform
Topics
Cost

Filtered by tag

Ask the library about workforce, startup, and technology trends

Answers come only from resources in this hub, with the sources listed underneath.

8 resources

FreeArticle
Technology & Ethics

"Alignment Engineering" vs. "Misalignment Science"

LessWrong post that separates AI alignment research into two categories, engineering work to make systems behave and scientific work to understand misalignment, in the context of debate over whether some alignment research is net negative.

#ai-alignment#ai-safety#lesswrong
lesswrong.comAdded Oct 7, 20260 opens
FreeBlog
Technology & Ethics

My New Course at UT Austin: AI Alignment Theory

October 2026 post by computer scientist Scott Aaronson describing CS395T AI Alignment Theory, a new graduate course he teaches at UT Austin, with the course description and topics.

#ai-alignment#course#ut-austin
scottaaronson.blogAdded Oct 7, 20260 opens
FreeEssay
Technology & Ethics

A Three-Facet Framework for AI Alignment

Essay proposing three distinct facets of AI alignment: control (the system cannot cause unacceptably bad outcomes even if trying), intent alignment (the system tries to do what the user wants), and values alignment (the system refuses widely unacceptable actions). Uses a stock-trading agent scenario to show how the facets can conflict.

#ai-alignment#ai-safety#ai-ethics
gracekind.netAdded Sep 30, 20260 opens
FreeDocument
Technology & Ethics

Model Misalignment Reporting Framework (OpenAI)

OpenAI's free framework for how misaligned model behaviour should be reported and categorised — what counts as misalignment, who reports it, and what happens next.

From the site: OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.

Why I recommend it: Primary source on how a major lab defines and handles its own model failures — useful, but it is the lab grading itself.

#ai#ai ethics#ai-alignment#ai-safety#ethics#governance#model-evaluation#openai#regulation#reporting#research#transparency
OpenAIAdded Sep 17, 20260 opens
FreeWebsite
Technology & Ethics

OpenAI Alignment — research and releases

OpenAI's free alignment research hub, including reports documenting how its own models fail.

From the site: Research on aligning AI with human values and intent, and reports documenting model failures.

Why I recommend it: A lab publishing on its own safety work — valuable primary material, but not an independent audit.

#ai#ai ethics#ai-alignment#ai-safety#ethics#llm#model-evaluation#openai#reports#research#transparency
alignment.openai.comAdded Sep 17, 20260 opens
FreeArticle
Technology & Ethics

What Is RLCD? Reinforcement Learning from Contrast Distillation

A plain-language explainer on RLCD, a way of aligning language models by learning from contrasting outputs rather than human ratings alone.

From the site: RLCD is a method developed to adjust language models to human preferences without using human feedback data. This approach aims to address…

Why I recommend it: Good background reading if you want to understand how the models you use are actually steered.

#ai#ai ethics#ai-alignment#ai-safety#explainers#llm#machine-learning#model-training#reinforcement-learning#research#rlhf
MediumAdded Sep 17, 20260 opens
FreeCommunity
Technology & Ethics

AI Alignment Forum

A community blog where researchers publish and debate technical AI alignment work, free and open to read.

From the site: A community blog devoted to technical AI alignment research

Why I recommend it: Dense reading, but this is where a lot of safety research is argued out in public before it reaches papers.

#academic#ai#ai ethics#ai-alignment#ai-ethics#ai-safety#community#free-resource#governance#machine-learning#regulation#research
alignmentforum.orgAdded Sep 17, 20260 opens