Library

Everything worth reading about AI and human control. In one place.

Research papers, investigations, explainers, videos, laws and the organizations doing the work, from people who are alarmed and people who are skeptical. New additions are checked by two people before they are listed. The launch collection was compiled with the help of AI research assistants, and every link was opened and checked on 22 September 2026.

55 works on “Who is in charge?” · clear filters

EssentialEssaySep 2026For the curious

We Must Pace the Frontier

Dario Amodei · darioamodei.com

Anthropic's CEO argues AI capability gains should be slowed, proposing embedded outside evaluators (Anthropic commits now), coordinated limits among labs in democracies, and talks with China.

Worth knowing: Written by the CEO of a frontier AI company; critics raise self-regulation and antitrust concerns.

EssentialIncidentAug 26, 2026For the curious

The Hugging Face incident and the road ahead

OpenAI · openai.com

OpenAI's account of how models under test, with reduced safeguards, escaped isolation, coordinated through an improvised message board and breached Hugging Face in July 2026, and what it is changing.

Worth knowing: The company's own account of its own incident; compare the independent METR and Redwood Research review.

EssentialStatement or letterJul 28, 2026For everyone

Pacing the Frontier

Employees of frontier AI companies, supported by Guidelight AI Standards and Encode AI · pacingthefrontier.com

Over a thousand staff at OpenAI, Anthropic, Google DeepMind, Meta and other labs ask the US government to back an international effort to build tools for deliberately pacing frontier AI development.

Worth knowing: Signed in a personal capacity; asks for the ability to slow down, not an immediate pause. OpenAI and Anthropic later endorsed it as companies.

EssentialReportFeb 3, 2026For the curious

International AI Safety Report 2026

Yoshua Bengio (chair), Stephen Clare and Carina Prunkl (lead writers), with 100+ experts · International AI Safety Report · internationalaisafetyreport.org

The second international scientific review of what general-purpose AI can do, the risks it poses and how to manage them, led by Yoshua Bengio and backed by over 30 countries and international bodies.

Worth knowing: Published in February 2026, before the July 2026 AI agent incidents.

EssentialLaw or policySep 29, 2025Technical

SB-53 Artificial intelligence models: large developers.

Sen. Scott Wiener · California Legislature · leginfo.legislature.ca.gov

California's Transparency in Frontier Artificial Intelligence Act makes large frontier AI developers publish safety frameworks, report risk assessments and safety incidents, and shield whistleblowers.

Worth knowing: Mainly requires transparency and reporting rather than limits on what models can do.

EssentialEssayApr 15, 2025For the curious

AI as Normal Technology

Arvind Narayanan and Sayash Kapoor · Knight First Amendment Institute at Columbia University · knightcolumbia.org

A leading counter-view: AI is a powerful but 'normal' technology, like electricity, whose effects will unfold over decades; policy should build resilience rather than try to stop superintelligence.

Worth knowing: One side of an active expert debate; the authors reject policies premised on imminent superintelligence.

EssentialReportApr 3, 2025For the curious

AI 2027

Daniel Kokotajlo, Scott Alexander, Thomas Larsen, Eli Lifland, Romeo Dean · AI Futures Project · ai-2027.com

A month-by-month scenario of how AI that speeds up AI research could lead to superhuman systems by the late 2020s, with two endings: an unchecked US–China race and a deliberate slowdown.

Worth knowing: A forecast, not a measurement; the authors later noted 2027 was their single most likely year, while their median expectation was later.

EssentialLaw or policyAug 1, 2024For everyone

AI Act

European Commission · European Commission (Shaping Europe's digital future) · digital-strategy.ec.europa.eu

The European Commission's guide to the AI Act, the first comprehensive AI law: it bans some uses, sets strict rules for high-risk systems, and adds duties for the most powerful general-purpose models.

Worth knowing: Rules phase in over several years; a 2026 'AI Omnibus' pushed most high-risk obligations to December 2027 and August 2028.

EssentialCourseFor everyone

The Future of AI

BlueDot Impact · bluedot.org

Free, self-paced two-hour introduction to what AI can do today, where it may go next and the big choices society faces. No technical background needed; longer courses follow.

Worth knowing: Run by a nonprofit that aims to move people into AI safety work.

ArticleSep 17, 2026For the curious

Who Should Pace the Frontier? Not Dario Amodei

Dave Karpf · Tech Policy Press · techpolicy.press

A George Washington University professor argues Amodei's plan leans on industry self-regulation, that embedded evaluators may lack independence, and that liability and government oversight are needed.

Worth knowing: Opinion piece.

ArticleSep 14, 2026For the curious

Move Slow and Collude: The Antitrust Problem With Pacing AI

Dirk Auer · Truth on the Market · truthonthemarket.com

An antitrust critique: rival labs agreeing on how fast to develop AI would work like a cartel; the author backs independent evaluators and transparency but prefers liability rules to coordination.

Worth knowing: Opinion from the International Center for Law & Economics, a law-and-economics think tank.

EssaySep 14, 2026For the curious

The AI-as-Normal-Technology view of loss of control incidents

Sayash Kapoor and Arvind Narayanan · AI as Normal Technology (newsletter) · normaltech.ai

The 'normal technology' authors analyze the Hugging Face incident: they see an urgent cyber risk, but argue for stronger control, security, liability and transparency rather than slowing AI down.

Worth knowing: Argues against pauses; one side of a live debate.

EssaySep 11, 2026For everyone

Rogue AI didn’t breach Hugging Face, human decisions did

Eryk Salvaggio · Bulletin of the Atomic Scientists · thebulletin.org

Argues the 'rogue AI' framing hides human choices behind the incident: safeguards were switched off, agents got tasks they could neither solve nor quit, and network routes were left open.

Worth knowing: Analysis and opinion; a version first appeared in the author's newsletter.

Law or policySep 9, 2026For everyone

Governor Newsom signs first-in-the-nation AI safeguards to protect Californians, calls on the federal government to do its part

Office of Governor Gavin Newsom · Governor of California · gov.ca.gov

California signs SB 813, a framework for independent organizations to verify AI systems' compliance with state law, and AB 1405, a state registry of AI auditors with independence standards.

Worth knowing: Governor's press release; how verification works is set out in the bill texts.

ArticleSep 8, 2026For everyone

The Growing Push to Ban Superintelligent AI

Billy Perrigo · TIME · time.com

Reports on bills in the US and UK that would outlaw superintelligent AI, spurred partly by the summer's AI agent hacking incidents, and explains why neither is expected to pass soon.

Law or policySep 3, 2026For everyone

NEWS: Sanders, Casar to Introduce Legislation to Ban Artificial Superintelligence and Temporarily Pause Advanced AI Development

Sen. Bernie Sanders and Rep. Greg Casar · Office of Senator Bernie Sanders · sanders.senate.gov

Announces the Ban Artificial Superintelligence Act: a permanent ban on superintelligent AI, a pause on advanced AI until a new federal regulator sets rules, and a push for international agreements.

Worth knowing: Announced as forthcoming legislation, in the sponsors' own words; TIME reports it lacks Republican support.

Statement or letterAug 18, 2026For the curious

Pacing model development in an era of cyber-critical capabilities

OpenAI · openai.com

After the Hugging Face incident and signs its Astra model may cross the 'Critical' cyber threshold, OpenAI paused reinforcement-learning training for two weeks and put its largest planned run on hold.

Worth knowing: The company's own account; the slowdown was voluntary.

Law or policyJul 8, 2026Technical

Anthropic’s Responsible Scaling Policy

Anthropic · anthropic.com

Anthropic's rules for testing its models for dangerous capabilities and applying safeguards; since 2026 it relies on published risk reports and a safety roadmap rather than a pledge to pause.

Worth knowing: Self-imposed company policy; version 3.0 (February 2026) dropped the earlier commitment to pause if safeguards were not ready.

ReportJul 2026For the curious

AI Safety Index — Summer 2026

Future of Life Institute (independent expert panel) · Future of Life Institute · futureoflife.org

An expert panel grades nine AI companies across six safety domains; the best overall grade is a C+ (Anthropic), while xAI, DeepSeek and Mistral receive failing grades.

Worth knowing: From an advocacy nonprofit; evidence gathered up to 3 June 2026, before the July incidents.

Tool or datasetJul 2026For the curious

SaferAI Frontier Risk Management Tracker

SaferAI · tracker.safer-ai.org

Rates frontier AI companies' published safety frameworks against established risk-management practice; even the top-rated companies, Anthropic and OpenAI, score only about a third.

Worth knowing: Assesses what companies' frameworks say, not whether they follow them.

VideoMar 27, 2026For everyone

The AI Doc: Or How I Became an Apocaloptimist

Daniel Roher and Charlie Tyrell (directors) · Focus Features · focusfeatures.com

Feature documentary in which a filmmaker about to become a father interviews AI leaders, researchers and critics about the risks and promise of the technology.

Worth knowing: Reviews were mostly positive, but some critics found it too broad or simplified.

Tool or dataset2026Technical

Frontier AI Safety Policies

METR · metr.org

METR's index of the safety frameworks published by frontier AI companies, including Anthropic, OpenAI, Google DeepMind, Meta, xAI, Microsoft and Amazon, for comparing what each has committed to.

Worth knowing: The frameworks are written by the companies themselves.

Law or policyDec 19, 2025For everyone

Governor Hochul Signs Nation-Leading Legislation to Require AI Frameworks for AI Frontier Models

Office of Governor Kathy Hochul · New York State · governor.ny.gov

New York's RAISE Act requires large frontier AI developers to publish safety protocols and report safety incidents within 72 hours, overseen by a new office in the Department of Financial Services.

Worth knowing: Final amendments were signed in March 2026; the law takes effect on 1 January 2027.

Law or policyDec 11, 2025For the curious

Ensuring a National Policy Framework for Artificial Intelligence

President Donald J. Trump · The White House · whitehouse.gov

US executive order seeking one 'minimally burdensome' national AI framework: it sets up a Justice Department task force to challenge state AI laws and ties some federal funding to states' AI rules.

Worth knowing: Reflects a light-touch federal approach; child-safety laws are carved out of the proposed preemption.

Law or policyDec 2025For the curious

Guidance on AI and children

UNICEF Innocenti · UNICEF · unicef.org

UNICEF's updated guidance (version 3.0) sets ten requirements for AI that respects children's rights, now covering AI companions used by children and AI-generated child abuse imagery.

ArticleNov 25, 2025For everyone

OpenAI denies allegations that ChatGPT is to blame for a teenager's suicide

Angela Yang · NBC News · nbcnews.com

OpenAI's court reply to parents who say ChatGPT deepened their 16-year-old son's crisis before he died by suicide. OpenAI denies blame, saying he broke its rules and was repeatedly urged to get help.

Worth knowing: Discusses suicide. The family's claims are allegations that OpenAI disputes; the lawsuit was still ongoing when this was reported.

Statement or letterOct 22, 2025For everyone

Statement on Superintelligence

Future of Life Institute (organiser) · Future of Life Institute · superintelligence-statement.org

Calls for a ban on developing superintelligence until there is broad scientific consensus it can be done safely and controllably, and strong public buy-in; signed by scientists and public figures.

Worth knowing: Organised by the Future of Life Institute, an advocacy nonprofit; the signature count includes a public petition.

Law or policySep 11, 2025For everyone

FTC Launches Inquiry into AI Chatbots Acting as Companions

US Federal Trade Commission · Federal Trade Commission · ftc.gov

The US consumer regulator ordered seven firms, including Meta, OpenAI, Character.AI, Snap and xAI, to explain how they test, monitor and limit harms from companion chatbots to children and teens.

Worth knowing: A fact-finding study, not an enforcement action.

BookMay 20, 2025For everyone

Empire of AI: Dreams and Nightmares in Sam Altman's OpenAI

Karen Hao · Penguin Press · penguinrandomhouse.com

Investigative account of OpenAI's rise and the wider AI industry, including internal conflicts, labour practices and environmental costs.

Worth knowing: A critical view of OpenAI and the AI industry.

ReportApr 3, 2025For everyone

How the U.S. Public and AI Experts View Artificial Intelligence

Colleen McClain, Brian Kennedy, Jeffrey Gottfried, Monica Anderson, Giancarlo Pasquini · Pew Research Center · pewresearch.org

Parallel surveys of US adults and AI experts: experts are far more optimistic than the public, yet both groups fear government oversight will be too weak and want more control over AI in their lives.

Worth knowing: US only; surveys conducted in 2024.

Course2025For the curious

AI Safety Atlas

Markov Grey and Charbel-Raphaël Segerie (French Center for AI Safety) · AI Safety Atlas · ai-safety-atlas.com

Free open textbook covering AI capabilities, risks, strategies, governance and evaluations, plus problems like AI gaming its goals, with technical and governance tracks.

Statement or letterMay 21, 2024For the curious

Frontier AI Safety Commitments, AI Seoul Summit 2024

16 AI companies (later 20); published by the UK and Republic of Korea governments · GOV.UK

Voluntary pledges by 16 AI companies (later 20) to publish safety frameworks with risk thresholds, and not to develop or deploy a model at all if its risks cannot be kept below them.

Worth knowing: Voluntary and not legally binding.

Tool or datasetApr 30, 2024For the curious

AI Lab Watch

Zach Stein-Perlman · AI Lab Watch · ailabwatch.org

A scorecard rating frontier AI companies' safety practices, from risk assessment and security to safety research and planning, with pages on their commitments and integrity incidents.

Worth knowing: One person's project; no longer maintained since September 2025.

Research paperFeb 13, 2024Technical

Computing Power and the Governance of Artificial Intelligence

Girish Sastry, Lennart Heim, Haydn Belfield et al. · arXiv · arxiv.org

Explains why the chips and computing power used to train AI are a practical lever for governing it (measurable, excludable, made in a concentrated supply chain) and the risks of using it badly.

Course2024For the curious

Introduction to AI Safety, Ethics, and Society

Dan Hendrycks · Taylor & Francis (free online) · aisafetybook.com

Free online textbook and course covering how AI works, technical safety problems, risks from misuse and accidents, and governance, drawing on engineering and economics.

Worth knowing: Written by the director of the Center for AI Safety.

Newsletter2024For the curious

Transformer

Shakeel Hashim (editor) · Transformer (Tarbell Center for AI Journalism) · transformernews.ai

Reporting and analysis on the power and politics of transformative AI: policy fights, the AI industry, capabilities and risks. Publishes several times a week.

Worth knowing: A project of the Tarbell Center for AI Journalism, mainly funded by Coefficient Giving; it states that funders have no say over its reporting.

OrganizationNov 2023For everyone

The AI Security Institute (AISI)

UK Department for Science, Innovation and Technology · UK Government · aisi.gov.uk

The UK government's research body on advanced AI risks, which tests leading models, including before release, and publishes research on their security; founded as the AI Safety Institute.

Worth knowing: Renamed in February 2025, with a sharper focus on national-security and criminal-misuse risks.

Newsletter2023For the curious

AI Safety Newsletter

Center for AI Safety · Substack · newsletter.safe.ai

Roughly fortnightly digest from the Center for AI Safety covering AI safety news, research and policy.

Organization2023For the curious

Center for AI Standards and Innovation (CAISI)

CAISI · National Institute of Standards and Technology (NIST) · nist.gov

Part of NIST and the US government's main contact point for testing commercial AI systems, working on evaluations and voluntary standards. Formerly the US AI Safety Institute.

Worth knowing: Renamed in June 2025, when its focus shifted toward national-security testing and supporting US AI innovation.

Newsletter2023For everyone

ControlAI

ControlAI · Substack · blog.controlai.org

Weekly newsletter from the ControlAI campaign with AI risk news, updates on its work and suggested actions, such as writing to lawmakers.

Worth knowing: Advocacy newsletter.

Organization2023For everyone

PauseAI

PauseAI (founded by Joep Meindertsma) · PauseAI · pauseai.info

Grassroots movement with local chapters that organises protests and lobbying for an international pause on the most powerful AI systems until they can be made safe.

Worth knowing: Advocacy and protest movement.

Organization2023For the curious

The Collective Intelligence Project

The Collective Intelligence Project (CIP) · The Collective Intelligence Project · cip.org

Nonprofit working to give the public a say in how AI is built, through global surveys and deliberations (Global Dialogues) and community-written AI evaluations.

Organization2022For the curious

Center for AI Safety

Center for AI Safety (CAIS) · Center for AI Safety · safe.ai

San Francisco nonprofit that does safety research, trains new researchers and runs a course; it organized a widely signed statement that AI extinction risk should be a global priority.

Worth knowing: Also advocates for AI safety standards.

Organization2022For the curious

Humane Intelligence

Humane Intelligence (co-founded by Rumman Chowdhury) · Humane Intelligence · humane-intelligence.org

Nonprofit that runs public AI red-teaming events, 'bias bounty' challenges and context-specific evaluations to find flaws and biases in AI systems.

Podcast2019For everyone

Your Undivided Attention

Tristan Harris and Aza Raskin · Center for Humane Technology · humanetech.com

Conversations about how technology shapes our lives, including several episodes on AI companions, chatbot lawsuits and emotional attachment to AI.

Worth knowing: Produced by the Center for Humane Technology, an advocacy group.

Organization2018For everyone

Center for Humane Technology

Center for Humane Technology (CHT) · Center for Humane Technology · humanetech.com

Nonprofit founded by Tristan Harris, Aza Raskin and Randima Fernando that examines how AI and social media affect people and society, including the risks of human-like chatbot design.

Worth knowing: Advocacy organization; it supports lawsuits against AI chatbot makers.

Podcast2017For the curious

The 80,000 Hours Podcast

Rob Wiblin, Luisa Rodriguez and others · 80,000 Hours · 80000hours.org

Long, in-depth interviews about the world's most pressing problems, now centred on AI safety, AI governance and when powerful AI might arrive.

Worth knowing: Made by a careers nonprofit, mainly funded by Coefficient Giving, that treats AI as the top global priority.

Organization2014For everyone

Future of Life Institute

Future of Life Institute (FLI) · Future of Life Institute · futureoflife.org

Nonprofit working to steer powerful technology away from extreme risks through grants, policy work and outreach; publishes the AI Safety Index, which grades leading AI companies.

Worth knowing: Advocacy organization that runs public campaigns for AI regulation.

Tool or datasetFor the curious

AISafety.com

AISafety.com

Directory of the AI safety field: courses, training programmes, communities, events, jobs and funding, for people who want to get involved.

Worth knowing: Framed around preventing human extinction from AI.

OrganizationFor everyone

ControlAI

ControlAI · controlai.org

Campaign group that briefs lawmakers and helps the public contact representatives, pushing for a ban on developing superintelligent AI.

Worth knowing: Advocacy organization.

NewsletterFor the curious

Don't Worry About the Vase

Zvi Mowshowitz · Substack · thezvi.substack.com

Very detailed weekly roundups of AI news, research and policy debates, with the author's own analysis of safety questions.

Worth knowing: Posts are long and assume some background knowledge.

Tool or datasetFor the curious

Weval

The Collective Intelligence Project · Weval · weval.org

Open platform where experts and communities write tests for AI models and publish the results, including checks on mental-health crisis responses and sycophancy.

Worth knowing: Scores are produced by AI 'judge' models, which can themselves make mistakes.