Skip to content
Launchpad Library logo

arxiv

9 free resources on this topic. Everything here is free and hand-picked. You can also search within this topic.

FreePaper
Research & Papers

Message Passing Enables Efficient Reasoning

An arXiv paper showing how message-passing techniques let language models reason more efficiently, with the full text free to read and download.

Why I recommend it: A preprint — not yet peer-reviewed, so treat the results as preliminary until the work is replicated or published.

#research#reasoning#models#arxiv
arxiv.orgAdded Sep 25, 20260 opens
FreeResearch Paper
Technology & Ethics

The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It (arXiv)

Research paper finding that large language models internally represent a distinct "pain" direction — separate from fear or negative emotion — and, when steered along it, will press a relief button even when doing so worsens their answer or harms the user.

Why I recommend it: Free to read on arXiv (preprint, not yet peer-reviewed). Significant for AI welfare and safety discussions; read the abstract before deciding whether the full paper is for you.

#ai ethics#ai-safety#ai-welfare#arxiv#interpretability#regulation#research
arxiv.orgAdded Sep 22, 20260 opens
FreeDocument
AI & Assistive Tools

Concrete Problems in AI Safety

The 2016 paper that framed AI safety as a set of specific engineering problems — side effects, reward hacking, unsafe exploration — rather than a philosophical worry. Free on arXiv.

From the site: Rapid progress in machine learning and artificial intelligence (AI) has brought increasing attention to the potential impacts of AI technologies on society. In this paper we discuss one such potential impact: the problem of accidents in machine learning systems, defined as unintended and harmful behavior that may emer…

Why I recommend it: Start here if the safety conversation sounds abstract. It is plain about what can go wrong and why, and almost everything since cites it.

#ai ethics#ai-risk#ai-safety#alignment#arxiv#foundational#free#machine-learning#reading#research#technology-and-ethics
arXiv.orgAdded Sep 17, 20260 opens
FreeDocument
AI & Assistive Tools

Managing Extreme AI Risks Amid Rapid Progress

A short consensus paper from Geoffrey Hinton, Yoshua Bengio and two dozen other researchers on the risks they consider serious and the governance they think is needed. Free on arXiv.

From the site: Artificial Intelligence (AI) is progressing rapidly, and companies are shifting their focus to developing generalist AI systems that can autonomously act and pursue goals. Increases in capabilities and autonomy may soon massively amplify AI's impact, with risks that include large-scale social harms, malicious uses, an…

Why I recommend it: The clearest statement of what the safety-concerned researchers actually agree on, signed rather than paraphrased.

#ai ethics#ai-policy#ai-risk#ai-safety#alignment#arxiv#free#governance#reading#regulation#research#technology-and-ethics
arXiv.orgAdded Sep 17, 20260 opens
FreeDocument
AI & Assistive Tools

Constitutional AI: Harmlessness from AI Feedback

Anthropic's paper describing how Claude is trained against a written set of principles instead of relying only on human ratings. Free on arXiv.

From the site: As AI systems become more capable, we would like to enlist their help to supervise other AIs. We experiment with methods for training a harmless AI assistant through self-improvement, without any human labels identifying harmful outputs. The only human oversight is provided through a list of rules or principles, and s…

Why I recommend it: Worth reading to see what "aligned" means in practice at one lab — and note it comes from the company selling the model.

#ai ethics#ai-risk#ai-safety#alignment#anthropic#arxiv#free#reading#research#technology-and-ethics#training
arXiv.orgAdded Sep 17, 20260 opens
FreeDocument
Technology & Ethics

Dream-RSI: Recursive Self-Improvement through Evolving Worlds

A free arXiv preprint on recursive self-improvement in AI agents trained inside evolving simulated worlds.

Why I recommend it: Technical, and central to the safety debate about systems that improve themselves.

#agents#ai#ai ethics#ai-risk#ai-safety#arxiv#machine-learning#preprint#recursive-self-improvement#research#simulation
arxiv.orgAdded Sep 17, 20260 opens
FreeWebsite
AI & Assistive Tools

Emergent Mind: Frontier Research Explorer

A free explorer for new arXiv research with plain-language paper summaries, topic pages and video overviews, so you can follow AI research without reading raw papers.

From the site: Your first stop to discover and learn about new arXiv research. Detailed paper summaries, video overviews, and more — no prompting required.

Why I recommend it: The fastest way I know to keep up with AI research when you are not a researcher. Free to browse.

#ai#research#arxiv#paper-summaries#ai-research#learning#reading#machine-learning#explainers#free-tools
emergentmind.comAdded Sep 17, 20260 opens
FreeArticle
Technology & Ethics

Design Docs Are All You Need (SMART research paper)

A research paper describing a software library whose repository holds almost no code: plain-language design documents are the durable artifact, and AI coding agents regenerate the implementation from those docs on every update.

Why I recommend it: The takeaway for non-engineers is bigger than the paper: clear written thinking is becoming the valuable skill, and the code is what gets generated from it.

#ai#ai ethics#article#arxiv#documentation#free#future-of-work#reading#research#software-engineering#tech-ethics#technology#writing
arxiv.orgAdded Sep 9, 20260 opens
FreeCase Study
Technology & Ethics

How AI Agents Reshape Knowledge Work: Autonomy, Efficiency, and Scope (working paper)

The underlying working paper by Jeremy Yang and co-authors, using Perplexity data to model tasks as discrete steps and compare fixed vs. marginal costs of chatbots versus autonomous agents.

Why I recommend it: If the HBS summary hooks you, go to the source. Skim the task-cost framework and use it to audit your own week: which tasks are high-step and repeatable? Those are the ones to hand to an agent first.

#ai#ai ethics#ai-agents#arxiv#automation#free#future-of-work#labor-economics#perplexity#research#research-paper#task-automation#tech-ethics#technology
arXivAdded Aug 30, 20260 opens