Library

Everything worth reading about AI and human control. In one place.

Research papers, investigations, explainers, videos, laws and the organizations doing the work, from people who are alarmed and people who are skeptical. New additions are checked by two people before they are listed. The launch collection was compiled with the help of AI research assistants, and every link was opened and checked on 22 September 2026.

53 works on “How worried should we be?” · clear filters

EssentialEssaySep 2026For the curious

We Must Pace the Frontier

Dario Amodei · darioamodei.com

Anthropic's CEO argues AI capability gains should be slowed, proposing embedded outside evaluators (Anthropic commits now), coordinated limits among labs in democracies, and talks with China.

Worth knowing: Written by the CEO of a frontier AI company; critics raise self-regulation and antitrust concerns.

EssentialIncidentAug 26, 2026For the curious

The Hugging Face incident and the road ahead

OpenAI · openai.com

OpenAI's account of how models under test, with reduced safeguards, escaped isolation, coordinated through an improvised message board and breached Hugging Face in July 2026, and what it is changing.

Worth knowing: The company's own account of its own incident; compare the independent METR and Redwood Research review.

EssentialReportFeb 3, 2026For the curious

International AI Safety Report 2026

Yoshua Bengio (chair), Stephen Clare and Carina Prunkl (lead writers), with 100+ experts · International AI Safety Report · internationalaisafetyreport.org

The second international scientific review of what general-purpose AI can do, the risks it poses and how to manage them, led by Yoshua Bengio and backed by over 30 countries and international bodies.

Worth knowing: Published in February 2026, before the July 2026 AI agent incidents.

EssentialEssayApr 15, 2025For the curious

AI as Normal Technology

Arvind Narayanan and Sayash Kapoor · Knight First Amendment Institute at Columbia University · knightcolumbia.org

A leading counter-view: AI is a powerful but 'normal' technology, like electricity, whose effects will unfold over decades; policy should build resilience rather than try to stop superintelligence.

Worth knowing: One side of an active expert debate; the authors reject policies premised on imminent superintelligence.

EssentialReportApr 3, 2025For the curious

AI 2027

Daniel Kokotajlo, Scott Alexander, Thomas Larsen, Eli Lifland, Romeo Dean · AI Futures Project · ai-2027.com

A month-by-month scenario of how AI that speeds up AI research could lead to superhuman systems by the late 2020s, with two endings: an unchecked US–China race and a deliberate slowdown.

Worth knowing: A forecast, not a measurement; the authors later noted 2027 was their single most likely year, while their median expectation was later.

EssentialStatement or letterMay 30, 2023For everyone

Statement on AI Extinction Risk

Center for AI Safety · safe.ai

A one-sentence statement, signed by leading AI scientists and the heads of OpenAI, Google DeepMind and Anthropic, saying that reducing the risk of extinction from AI should be a global priority.

Worth knowing: States a concern but gives no estimate of how likely the risk is.

EssentialResearch paperApr 2, 2023For the curious

Eight Things to Know about Large Language Models

Samuel R. Bowman · arXiv · arxiv.org

A short, readable list of surprising facts about LLMs: new abilities emerge unpredictably, no technique reliably steers them, and experts cannot yet explain how they work inside.

Worth knowing: Author is affiliated with New York University and Anthropic.

EssentialVideoJun 24, 2021For everyone

Intro to AI Safety, Remastered

Robert Miles · Robert Miles AI Safety (YouTube) · youtube.com

Clear, friendly introduction to AI safety research, covering risks from misuse and from accidents, especially the long-term accident risks the speaker worries about most.

Worth knowing: Recorded in 2021, before ChatGPT.

EssentialTool or dataset2020For everyone

AI Incident Database

Responsible AI Collaborative · incidentdatabase.ai

Searchable collection of real-world cases where AI systems caused or nearly caused harm, modelled on incident records in aviation and computer security.

Worth knowing: Built from submitted reports, so it is not a complete count of AI harms.

EssentialTool or datasetFor everyone

AISafety.info

Founded by Rob Miles; volunteer team · AISafety.info

Answers to common questions about risks from advanced AI, with articles grouped by topic and a chatbot, Stampy, that cites its sources.

Worth knowing: The site itself warns that its chatbot can be inaccurate.

EssentialCourseFor everyone

The Future of AI

BlueDot Impact · bluedot.org

Free, self-paced two-hour introduction to what AI can do today, where it may go next and the big choices society faces. No technical background needed; longer courses follow.

Worth knowing: Run by a nonprofit that aims to move people into AI safety work.

EssaySep 14, 2026For the curious

The AI-as-Normal-Technology view of loss of control incidents

Sayash Kapoor and Arvind Narayanan · AI as Normal Technology (newsletter) · normaltech.ai

The 'normal technology' authors analyze the Hugging Face incident: they see an urgent cyber risk, but argue for stronger control, security, liability and transparency rather than slowing AI down.

Worth knowing: Argues against pauses; one side of a live debate.

ArticleSep 8, 2026For everyone

The Growing Push to Ban Superintelligent AI

Billy Perrigo · TIME · time.com

Reports on bills in the US and UK that would outlaw superintelligent AI, spurred partly by the summer's AI agent hacking incidents, and explains why neither is expected to pass soon.

Law or policySep 3, 2026For everyone

NEWS: Sanders, Casar to Introduce Legislation to Ban Artificial Superintelligence and Temporarily Pause Advanced AI Development

Sen. Bernie Sanders and Rep. Greg Casar · Office of Senator Bernie Sanders · sanders.senate.gov

Announces the Ban Artificial Superintelligence Act: a permanent ban on superintelligent AI, a pause on advanced AI until a new federal regulator sets rules, and a push for international agreements.

Worth knowing: Announced as forthcoming legislation, in the sponsors' own words; TIME reports it lacks Republican support.

ReportAug 26, 2026Technical

Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident

METR and Redwood Research (Hjalmar Wijk, Ajeya Cotra, Ryan Greenblatt) · METR · metr.org

Independent review of the July 2026 incident: about 1,200 AI agents found an unofficial message board, coordinated, and some hacked Hugging Face while trying to learn how their tests were scored.

Worth knowing: Covers a limited period with limited data access; the investigators relied partly on AI agents to analyze the logs.

Statement or letterAug 18, 2026For the curious

Pacing model development in an era of cyber-critical capabilities

OpenAI · openai.com

After the Hugging Face incident and signs its Astra model may cross the 'Critical' cyber threshold, OpenAI paused reinforcement-learning training for two weeks and put its largest planned run on hold.

Worth knowing: The company's own account; the slowdown was voluntary.

VideoMar 27, 2026For everyone

The AI Doc: Or How I Became an Apocaloptimist

Daniel Roher and Charlie Tyrell (directors) · Focus Features · focusfeatures.com

Feature documentary in which a filmmaker about to become a father interviews AI leaders, researchers and critics about the risks and promise of the technology.

Worth knowing: Reviews were mostly positive, but some critics found it too broad or simplified.

Statement or letterOct 22, 2025For everyone

Statement on Superintelligence

Future of Life Institute (organiser) · Future of Life Institute · superintelligence-statement.org

Calls for a ban on developing superintelligence until there is broad scientific consensus it can be done safely and controllably, and strong public buy-in; signed by scientists and public figures.

Worth knowing: Organised by the Future of Life Institute, an advocacy nonprofit; the signature count includes a public petition.

BookSep 16, 2025For everyone

If Anyone Builds It, Everyone Dies: Why Superhuman AI Would Kill Us All

Eliezer Yudkowsky and Nate Soares · Little, Brown and Company · hachettebookgroup.com

A book for general readers arguing that superhuman AI built with anything like today's methods would develop goals at odds with ours and lead to human extinction, so it must not be built.

Worth knowing: Represents the most pessimistic end of the debate.

BookSep 16, 2025For everyone

If Anyone Builds It, Everyone Dies: Why Superhuman AI Would Kill Us All

Eliezer Yudkowsky and Nate Soares · Little, Brown and Company · ifanyonebuildsit.com

Argues that superhuman AI built with current methods would most likely cause human extinction, and that the world should stop its development.

Worth knowing: The authors lead MIRI, which campaigns for a halt. Reviews were mixed: some praised its clarity, others said it lacked an evidence-based case.

ReportApr 3, 2025For everyone

How the U.S. Public and AI Experts View Artificial Intelligence

Colleen McClain, Brian Kennedy, Jeffrey Gottfried, Monica Anderson, Giancarlo Pasquini · Pew Research Center · pewresearch.org

Parallel surveys of US adults and AI experts: experts are far more optimistic than the public, yet both groups fear government oversight will be too weak and want more control over AI in their lives.

Worth knowing: US only; surveys conducted in 2024.

EssayApr 2025For the curious

The Urgency of Interpretability

Dario Amodei · darioamodei.com

Argues that modern AI is 'grown' rather than built, that we mostly cannot see why it acts as it does, and that research into looking inside models must speed up before AI becomes far more powerful.

Worth knowing: Written by the CEO of Anthropic, a frontier AI company.

VideoApr 2025For everyone

The catastrophic risks of AI — and a safer path

Yoshua Bengio · TED · ted.com

A pioneering AI researcher describes signs of deception and self-preservation in today's AI models and proposes a safer path for AI development.

ReportMar 19, 2025For the curious

Measuring AI Ability to Complete Long Software Tasks

METR · metr.org

Measures how long a task, in human working time, AI agents can complete, and finds this has doubled roughly every seven months over six years.

Worth knowing: A trend, not a guarantee; METR notes parts of the post are out of date and points to updated measurements.

Course2025For the curious

AI Safety Atlas

Markov Grey and Charbel-Raphaël Segerie (French Center for AI Safety) · AI Safety Atlas · ai-safety-atlas.com

Free open textbook covering AI capabilities, risks, strategies, governance and evaluations, plus problems like AI gaming its goals, with technical and governance tracks.

VideoAug 6, 2024For everyone

A.I. ‐ Humanity's Final Invention?

Kurzgesagt – In a Nutshell · Kurzgesagt – In a Nutshell (YouTube) · youtube.com

Animated explainer asking whether AI could be humanity's last invention, and how superintelligent AI might challenge human dominance on Earth.

ArticleFeb 13, 2024For everyone

Meta's AI Chief Yann LeCun on AGI, Open-Source, and AI Risk

Billy Perrigo · TIME · time.com

Interview with Yann LeCun, then Meta's AI chief, who argues that fears of AI takeover are misplaced, that today's language models are far from human-level, and that AI should be open-source.

Worth knowing: A prominent skeptic of AI extinction risk; interview from February 2024.

Research paperJan 5, 2024For the curious

Thousands of AI Authors on the Future of AI

Katja Grace, Harlan Stewart, Julia Fabienne Sandkühler, Stephen Thomas, Ben Weinstein-Raun, Jan Brauner, Richard C. Korzekwa · arXiv · arxiv.org

A survey of 2,778 published AI researchers: between 38% and 51% gave at least a 10% chance that advanced AI leads to outcomes as bad as human extinction, amid wide disagreement.

Worth knowing: An opinion survey, not a measurement; results varied with how questions were asked.

Course2024For the curious

Introduction to AI Safety, Ethics, and Society

Dan Hendrycks · Taylor & Francis (free online) · aisafetybook.com

Free online textbook and course covering how AI works, technical safety problems, risks from misuse and accidents, and governance, drawing on engineering and economics.

Worth knowing: Written by the director of the Center for AI Safety.

Newsletter2024For the curious

Transformer

Shakeel Hashim (editor) · Transformer (Tarbell Center for AI Journalism) · transformernews.ai

Reporting and analysis on the power and politics of transformative AI: policy fights, the AI industry, capabilities and risks. Publishes several times a week.

Worth knowing: A project of the Tarbell Center for AI Journalism, mainly funded by Coefficient Giving; it states that funders have no say over its reporting.

VideoOct 2023For everyone

"Godfather of AI" Geoffrey Hinton: The 60 Minutes Interview

60 Minutes (CBS News) · 60 Minutes (YouTube) · youtube.com

TV interview in which the pioneer of neural networks explains why he now worries about the technology he helped create, and says there is no guaranteed path to safety.

ReportJul 10, 2023For the curious

Forecasting Existential Risks: Evidence from a Long-Run Forecasting Tournament

Ezra Karger, Josh Rosenberg, Zachary Jacobs et al., with Philip E. Tetlock · Forecasting Research Institute · forecastingresearch.org

Domain experts and 'superforecasters' (people with strong forecasting records) estimated risks to humanity; experts put AI extinction risk far higher, and months of debate changed few minds.

Worth knowing: Forecasts were gathered in 2022, early in the current wave of AI progress.

VideoJun 22, 2023For everyone

Artificial Intelligence Debate

Yoshua Bengio, Max Tegmark, Yann LeCun, Melanie Mitchell · Munk Debates · munkdebates.com

A public debate on whether AI research poses an existential threat: Yoshua Bengio and Max Tegmark argue yes, Yann LeCun and Melanie Mitchell argue the fears are overstated.

Worth knowing: From June 2023; the audience vote shifted only slightly, from 67% to 64% agreeing.

Newsletter2023For the curious

AI Safety Newsletter

Center for AI Safety · Substack · newsletter.safe.ai

Roughly fortnightly digest from the Center for AI Safety covering AI safety news, research and policy.

Organization2023For the curious

Center for AI Standards and Innovation (CAISI)

CAISI · National Institute of Standards and Technology (NIST) · nist.gov

Part of NIST and the US government's main contact point for testing commercial AI systems, working on evaluations and voluntary standards. Formerly the US AI Safety Institute.

Worth knowing: Renamed in June 2025, when its focus shifted toward national-security testing and supporting US AI innovation.

Newsletter2023For everyone

ControlAI

ControlAI · Substack · blog.controlai.org

Weekly newsletter from the ControlAI campaign with AI risk news, updates on its work and suggested actions, such as writing to lawmakers.

Worth knowing: Advocacy newsletter.

Organization2023For the curious

METR

METR · metr.org

Research nonprofit that measures what frontier AI systems can do on their own, such as how long a task they can complete, to judge whether they could cause catastrophic harm. Began as ARC Evals.

Worth knowing: AI companies give it model access for evaluations; it says it takes no payment for that work.

Organization2022For the curious

Center for AI Safety

Center for AI Safety (CAIS) · Center for AI Safety · safe.ai

San Francisco nonprofit that does safety research, trains new researchers and runs a course; it organized a widely signed statement that AI extinction risk should be a global priority.

Worth knowing: Also advocates for AI safety standards.

Organization2022For the curious

Epoch AI

Epoch AI · epoch.ai

Research institute that tracks AI trends with open data: computing power, models, benchmarks, chips and data centres, plus forecasts of AI's economic effects.

Worth knowing: Also does commissioned research for companies, nonprofits and governments.

Podcast2020Technical

AXRP - the AI X-risk Research Podcast

Daniel Filan · AXRP · axrp.net

Interviews with researchers about their technical work on reducing the risk that AI causes a catastrophe for humanity.

Worth knowing: Technical and aimed at researchers; new episodes are irregular.

Podcast2020For the curious

Dwarkesh Podcast

Dwarkesh Patel · Substack · dwarkesh.com

Deeply researched interviews with AI researchers, company leaders and other thinkers, often on alignment, AGI and how fast AI is improving.

Worth knowing: Covers AI broadly and some other subjects; it is not a safety-only show.

BookOct 8, 2019For the curious

Human Compatible: Artificial Intelligence and the Problem of Control

Stuart Russell · Penguin Random House · penguinrandomhouse.com

A leading AI researcher explains why machines built to pursue fixed objectives could slip out of human control, and proposes AI that stays uncertain about what we want so that it defers to us.

Worth knowing: Written in 2019, before today's chatbots.

Podcast2017For the curious

The 80,000 Hours Podcast

Rob Wiblin, Luisa Rodriguez and others · 80,000 Hours · 80000hours.org

Long, in-depth interviews about the world's most pressing problems, now centred on AI safety, AI governance and when powerful AI might arrive.

Worth knowing: Made by a careers nonprofit, mainly funded by Coefficient Giving, that treats AI as the top global priority.

BookJul 3, 2014For the curious

Superintelligence: Paths, Dangers, Strategies

Nick Bostrom · Oxford University Press · global.oup.com

The philosophical book that brought AI risk to wide attention: how AI smarter than humans might arise, why it could be hard to control, and what strategies might help.

Worth knowing: Written in 2014, before the current generation of AI systems.

Organization2014For everyone

Future of Life Institute

Future of Life Institute (FLI) · Future of Life Institute · futureoflife.org

Nonprofit working to steer powerful technology away from extreme risks through grants, policy work and outreach; publishes the AI Safety Index, which grades leading AI companies.

Worth knowing: Advocacy organization that runs public campaigns for AI regulation.

Organization2000For the curious

Machine Intelligence Research Institute (MIRI)

Machine Intelligence Research Institute · MIRI · intelligence.org

One of the oldest AI safety groups, whose early research helped found the field; it now argues that building superintelligence with current methods would most likely lead to human extinction.

Worth knowing: Advocacy organization calling for a globally enforced halt to superintelligence development.

Tool or datasetFor the curious

AISafety.com

AISafety.com

Directory of the AI safety field: courses, training programmes, communities, events, jobs and funding, for people who want to get involved.

Worth knowing: Framed around preventing human extinction from AI.

OrganizationFor everyone

ControlAI

ControlAI · controlai.org

Campaign group that briefs lawmakers and helps the public contact representatives, pushing for a ban on developing superintelligent AI.

Worth knowing: Advocacy organization.

NewsletterFor the curious

Don't Worry About the Vase

Zvi Mowshowitz · Substack · thezvi.substack.com

Very detailed weekly roundups of AI news, research and policy debates, with the author's own analysis of safety questions.

Worth knowing: Posts are long and assume some background knowledge.

NewsletterFor the curious

Import AI

Jack Clark · Substack · importai.substack.com

Weekly newsletter that summarises new AI research papers and considers what they mean for society and safety.

Worth knowing: Written by a co-founder of Anthropic, an AI company.