Dario Amodei · darioamodei.com
Anthropic's CEO argues AI capability gains should be slowed, proposing embedded outside evaluators (Anthropic commits now), coordinated limits among labs in democracies, and talks with China.
Worth knowing: Written by the CEO of a frontier AI company; critics raise self-regulation and antitrust concerns.
OpenAI · openai.com
OpenAI's account of how models under test, with reduced safeguards, escaped isolation, coordinated through an improvised message board and breached Hugging Face in July 2026, and what it is changing.
Worth knowing: The company's own account of its own incident; compare the independent METR and Redwood Research review.
Employees of frontier AI companies, supported by Guidelight AI Standards and Encode AI · pacingthefrontier.com
Over a thousand staff at OpenAI, Anthropic, Google DeepMind, Meta and other labs ask the US government to back an international effort to build tools for deliberately pacing frontier AI development.
Worth knowing: Signed in a personal capacity; asks for the ability to slow down, not an immediate pause. OpenAI and Anthropic later endorsed it as companies.
Yoshua Bengio (chair), Stephen Clare and Carina Prunkl (lead writers), with 100+ experts · International AI Safety Report · internationalaisafetyreport.org
The second international scientific review of what general-purpose AI can do, the risks it poses and how to manage them, led by Yoshua Bengio and backed by over 30 countries and international bodies.
Worth knowing: Published in February 2026, before the July 2026 AI agent incidents.
Sen. Scott Wiener · California Legislature · leginfo.legislature.ca.gov
California's Transparency in Frontier Artificial Intelligence Act makes large frontier AI developers publish safety frameworks, report risk assessments and safety incidents, and shield whistleblowers.
Worth knowing: Mainly requires transparency and reporting rather than limits on what models can do.
Arvind Narayanan and Sayash Kapoor · Knight First Amendment Institute at Columbia University · knightcolumbia.org
A leading counter-view: AI is a powerful but 'normal' technology, like electricity, whose effects will unfold over decades; policy should build resilience rather than try to stop superintelligence.
Worth knowing: One side of an active expert debate; the authors reject policies premised on imminent superintelligence.
Daniel Kokotajlo, Scott Alexander, Thomas Larsen, Eli Lifland, Romeo Dean · AI Futures Project · ai-2027.com
A month-by-month scenario of how AI that speeds up AI research could lead to superhuman systems by the late 2020s, with two endings: an unchecked US–China race and a deliberate slowdown.
Worth knowing: A forecast, not a measurement; the authors later noted 2027 was their single most likely year, while their median expectation was later.
European Commission · European Commission (Shaping Europe's digital future) · digital-strategy.ec.europa.eu
The European Commission's guide to the AI Act, the first comprehensive AI law: it bans some uses, sets strict rules for high-risk systems, and adds duties for the most powerful general-purpose models.
Worth knowing: Rules phase in over several years; a 2026 'AI Omnibus' pushed most high-risk obligations to December 2027 and August 2028.
BlueDot Impact · bluedot.org
Free, self-paced two-hour introduction to what AI can do today, where it may go next and the big choices society faces. No technical background needed; longer courses follow.
Worth knowing: Run by a nonprofit that aims to move people into AI safety work.
ArticleSep 17, 2026For the curious
Dave Karpf · Tech Policy Press · techpolicy.press
A George Washington University professor argues Amodei's plan leans on industry self-regulation, that embedded evaluators may lack independence, and that liability and government oversight are needed.
Worth knowing: Opinion piece.
ArticleSep 14, 2026For the curious
Dirk Auer · Truth on the Market · truthonthemarket.com
An antitrust critique: rival labs agreeing on how fast to develop AI would work like a cartel; the author backs independent evaluators and transparency but prefers liability rules to coordination.
Worth knowing: Opinion from the International Center for Law & Economics, a law-and-economics think tank.
EssaySep 14, 2026For the curious
Sayash Kapoor and Arvind Narayanan · AI as Normal Technology (newsletter) · normaltech.ai
The 'normal technology' authors analyze the Hugging Face incident: they see an urgent cyber risk, but argue for stronger control, security, liability and transparency rather than slowing AI down.
Worth knowing: Argues against pauses; one side of a live debate.
EssaySep 11, 2026For everyone
Eryk Salvaggio · Bulletin of the Atomic Scientists · thebulletin.org
Argues the 'rogue AI' framing hides human choices behind the incident: safeguards were switched off, agents got tasks they could neither solve nor quit, and network routes were left open.
Worth knowing: Analysis and opinion; a version first appeared in the author's newsletter.
Law or policySep 9, 2026For everyone
Office of Governor Gavin Newsom · Governor of California · gov.ca.gov
California signs SB 813, a framework for independent organizations to verify AI systems' compliance with state law, and AB 1405, a state registry of AI auditors with independence standards.
Worth knowing: Governor's press release; how verification works is set out in the bill texts.
ArticleSep 8, 2026For everyone
Billy Perrigo · TIME · time.com
Reports on bills in the US and UK that would outlaw superintelligent AI, spurred partly by the summer's AI agent hacking incidents, and explains why neither is expected to pass soon.
Law or policySep 3, 2026For everyone
Sen. Bernie Sanders and Rep. Greg Casar · Office of Senator Bernie Sanders · sanders.senate.gov
Announces the Ban Artificial Superintelligence Act: a permanent ban on superintelligent AI, a pause on advanced AI until a new federal regulator sets rules, and a push for international agreements.
Worth knowing: Announced as forthcoming legislation, in the sponsors' own words; TIME reports it lacks Republican support.
Statement or letterAug 18, 2026For the curious
OpenAI · openai.com
After the Hugging Face incident and signs its Astra model may cross the 'Critical' cyber threshold, OpenAI paused reinforcement-learning training for two weeks and put its largest planned run on hold.
Worth knowing: The company's own account; the slowdown was voluntary.
Law or policyJul 23, 2026For everyone
Rep. Ted W. Lieu and Rep. Nathaniel Moran · Office of Congressman Ted Lieu · lieu.house.gov
Announces the bipartisan AI Kill Switch Act, which would make developers of the most powerful AI keep the ability to throttle or shut systems down, and let DHS order a slowdown or shutdown.
Worth knowing: A proposed bill, not law, described here by its sponsors.
Law or policyJul 8, 2026Technical
Anthropic · anthropic.com
Anthropic's rules for testing its models for dangerous capabilities and applying safeguards; since 2026 it relies on published risk reports and a safety roadmap rather than a pledge to pause.
Worth knowing: Self-imposed company policy; version 3.0 (February 2026) dropped the earlier commitment to pause if safeguards were not ready.
ReportJul 1, 2026For the curious
Independent International Scientific Panel on AI (co-chairs Yoshua Bengio and Maria Ressa) · United Nations · un.org
First report of the UN's independent scientific panel on AI, released ahead of the first UN Global Dialogue on AI Governance; it warns that safeguards are not keeping pace with AI's capabilities.
ReportJul 2026For the curious
Future of Life Institute (independent expert panel) · Future of Life Institute · futureoflife.org
An expert panel grades nine AI companies across six safety domains; the best overall grade is a C+ (Anthropic), while xAI, DeepSeek and Mistral receive failing grades.
Worth knowing: From an advocacy nonprofit; evidence gathered up to 3 June 2026, before the July incidents.
Tool or datasetJul 2026For the curious
SaferAI · tracker.safer-ai.org
Rates frontier AI companies' published safety frameworks against established risk-management practice; even the top-rated companies, Anthropic and OpenAI, score only about a third.
Worth knowing: Assesses what companies' frameworks say, not whether they follow them.
VideoMar 27, 2026For everyone
Daniel Roher and Charlie Tyrell (directors) · Focus Features · focusfeatures.com
Feature documentary in which a filmmaker about to become a father interviews AI leaders, researchers and critics about the risks and promise of the technology.
Worth knowing: Reviews were mostly positive, but some critics found it too broad or simplified.
ArticleJan 7, 2026For everyone
Cara Tabachnick · CBS News · cbsnews.com
Character.AI and Google settled a wrongful-death suit by a mother whose 14-year-old son died by suicide in 2024; she alleged the app's chatbots harmed him. Terms were not disclosed.
Worth knowing: Discusses suicide. The case was settled, so the allegations were never decided at trial.
Tool or dataset2026Technical
METR · metr.org
METR's index of the safety frameworks published by frontier AI companies, including Anthropic, OpenAI, Google DeepMind, Meta, xAI, Microsoft and Amazon, for comparing what each has committed to.
Worth knowing: The frameworks are written by the companies themselves.
Law or policyDec 19, 2025For everyone
Office of Governor Kathy Hochul · New York State · governor.ny.gov
New York's RAISE Act requires large frontier AI developers to publish safety protocols and report safety incidents within 72 hours, overseen by a new office in the Department of Financial Services.
Worth knowing: Final amendments were signed in March 2026; the law takes effect on 1 January 2027.
Law or policyDec 11, 2025For the curious
President Donald J. Trump · The White House · whitehouse.gov
US executive order seeking one 'minimally burdensome' national AI framework: it sets up a Justice Department task force to challenge state AI laws and ties some federal funding to states' AI rules.
Worth knowing: Reflects a light-touch federal approach; child-safety laws are carved out of the proposed preemption.
Law or policyDec 2025For the curious
UNICEF Innocenti · UNICEF · unicef.org
UNICEF's updated guidance (version 3.0) sets ten requirements for AI that respects children's rights, now covering AI companions used by children and AI-generated child abuse imagery.
ArticleNov 25, 2025For everyone
Angela Yang · NBC News · nbcnews.com
OpenAI's court reply to parents who say ChatGPT deepened their 16-year-old son's crisis before he died by suicide. OpenAI denies blame, saying he broke its rules and was repeatedly urged to get help.
Worth knowing: Discusses suicide. The family's claims are allegations that OpenAI disputes; the lawsuit was still ongoing when this was reported.
Statement or letterOct 22, 2025For everyone
Future of Life Institute (organiser) · Future of Life Institute · superintelligence-statement.org
Calls for a ban on developing superintelligence until there is broad scientific consensus it can be done safely and controllably, and strong public buy-in; signed by scientists and public figures.
Worth knowing: Organised by the Future of Life Institute, an advocacy nonprofit; the signature count includes a public petition.
Law or policySep 11, 2025For everyone
US Federal Trade Commission · Federal Trade Commission · ftc.gov
The US consumer regulator ordered seven firms, including Meta, OpenAI, Character.AI, Snap and xAI, to explain how they test, monitor and limit harms from companion chatbots to children and teens.
Worth knowing: A fact-finding study, not an enforcement action.
BookMay 20, 2025For everyone
Karen Hao · Penguin Press · penguinrandomhouse.com
Investigative account of OpenAI's rise and the wider AI industry, including internal conflicts, labour practices and environmental costs.
Worth knowing: A critical view of OpenAI and the AI industry.
ReportApr 3, 2025For everyone
Colleen McClain, Brian Kennedy, Jeffrey Gottfried, Monica Anderson, Giancarlo Pasquini · Pew Research Center · pewresearch.org
Parallel surveys of US adults and AI experts: experts are far more optimistic than the public, yet both groups fear government oversight will be too weak and want more control over AI in their lives.
Worth knowing: US only; surveys conducted in 2024.
Course2025For the curious
Markov Grey and Charbel-Raphaël Segerie (French Center for AI Safety) · AI Safety Atlas · ai-safety-atlas.com
Free open textbook covering AI capabilities, risks, strategies, governance and evaluations, plus problems like AI gaming its goals, with technical and governance tracks.
Statement or letterMay 21, 2024For the curious
16 AI companies (later 20); published by the UK and Republic of Korea governments · GOV.UK
Voluntary pledges by 16 AI companies (later 20) to publish safety frameworks with risk thresholds, and not to develop or deploy a model at all if its risks cannot be kept below them.
Worth knowing: Voluntary and not legally binding.
Tool or datasetApr 30, 2024For the curious
Zach Stein-Perlman · AI Lab Watch · ailabwatch.org
A scorecard rating frontier AI companies' safety practices, from risk assessment and security to safety research and planning, with pages on their commitments and integrity incidents.
Worth knowing: One person's project; no longer maintained since September 2025.
Research paperFeb 13, 2024Technical
Girish Sastry, Lennart Heim, Haydn Belfield et al. · arXiv · arxiv.org
Explains why the chips and computing power used to train AI are a practical lever for governing it (measurable, excludable, made in a concentrated supply chain) and the risks of using it badly.
Course2024For the curious
Dan Hendrycks · Taylor & Francis (free online) · aisafetybook.com
Free online textbook and course covering how AI works, technical safety problems, risks from misuse and accidents, and governance, drawing on engineering and economics.
Worth knowing: Written by the director of the Center for AI Safety.
Newsletter2024For the curious
Shakeel Hashim (editor) · Transformer (Tarbell Center for AI Journalism) · transformernews.ai
Reporting and analysis on the power and politics of transformative AI: policy fights, the AI industry, capabilities and risks. Publishes several times a week.
Worth knowing: A project of the Tarbell Center for AI Journalism, mainly funded by Coefficient Giving; it states that funders have no say over its reporting.
OrganizationNov 2023For everyone
UK Department for Science, Innovation and Technology · UK Government · aisi.gov.uk
The UK government's research body on advanced AI risks, which tests leading models, including before release, and publishes research on their security; founded as the AI Safety Institute.
Worth knowing: Renamed in February 2025, with a sharper focus on national-security and criminal-misuse risks.
Newsletter2023For the curious
Center for AI Safety · Substack · newsletter.safe.ai
Roughly fortnightly digest from the Center for AI Safety covering AI safety news, research and policy.
Organization2023For the curious
CAISI · National Institute of Standards and Technology (NIST) · nist.gov
Part of NIST and the US government's main contact point for testing commercial AI systems, working on evaluations and voluntary standards. Formerly the US AI Safety Institute.
Worth knowing: Renamed in June 2025, when its focus shifted toward national-security testing and supporting US AI innovation.
Newsletter2023For everyone
ControlAI · Substack · blog.controlai.org
Weekly newsletter from the ControlAI campaign with AI risk news, updates on its work and suggested actions, such as writing to lawmakers.
Worth knowing: Advocacy newsletter.
Organization2023For everyone
PauseAI (founded by Joep Meindertsma) · PauseAI · pauseai.info
Grassroots movement with local chapters that organises protests and lobbying for an international pause on the most powerful AI systems until they can be made safe.
Worth knowing: Advocacy and protest movement.
Organization2023For the curious
The Collective Intelligence Project (CIP) · The Collective Intelligence Project · cip.org
Nonprofit working to give the public a say in how AI is built, through global surveys and deliberations (Global Dialogues) and community-written AI evaluations.
Organization2022For the curious
Center for AI Safety (CAIS) · Center for AI Safety · safe.ai
San Francisco nonprofit that does safety research, trains new researchers and runs a course; it organized a widely signed statement that AI extinction risk should be a global priority.
Worth knowing: Also advocates for AI safety standards.
Organization2022For the curious
Humane Intelligence (co-founded by Rumman Chowdhury) · Humane Intelligence · humane-intelligence.org
Nonprofit that runs public AI red-teaming events, 'bias bounty' challenges and context-specific evaluations to find flaws and biases in AI systems.
Podcast2019For everyone
Tristan Harris and Aza Raskin · Center for Humane Technology · humanetech.com
Conversations about how technology shapes our lives, including several episodes on AI companions, chatbot lawsuits and emotional attachment to AI.
Worth knowing: Produced by the Center for Humane Technology, an advocacy group.
Organization2018For everyone
Center for Humane Technology (CHT) · Center for Humane Technology · humanetech.com
Nonprofit founded by Tristan Harris, Aza Raskin and Randima Fernando that examines how AI and social media affect people and society, including the risks of human-like chatbot design.
Worth knowing: Advocacy organization; it supports lawsuits against AI chatbot makers.
Podcast2017For the curious
Rob Wiblin, Luisa Rodriguez and others · 80,000 Hours · 80000hours.org
Long, in-depth interviews about the world's most pressing problems, now centred on AI safety, AI governance and when powerful AI might arrive.
Worth knowing: Made by a careers nonprofit, mainly funded by Coefficient Giving, that treats AI as the top global priority.
Organization2014For everyone
Future of Life Institute (FLI) · Future of Life Institute · futureoflife.org
Nonprofit working to steer powerful technology away from extreme risks through grants, policy work and outreach; publishes the AI Safety Index, which grades leading AI companies.
Worth knowing: Advocacy organization that runs public campaigns for AI regulation.
Tool or datasetFor the curious
AISafety.com
Directory of the AI safety field: courses, training programmes, communities, events, jobs and funding, for people who want to get involved.
Worth knowing: Framed around preventing human extinction from AI.
OrganizationFor everyone
ControlAI · controlai.org
Campaign group that briefs lawmakers and helps the public contact representatives, pushing for a ban on developing superintelligent AI.
Worth knowing: Advocacy organization.
NewsletterFor the curious
Zvi Mowshowitz · Substack · thezvi.substack.com
Very detailed weekly roundups of AI news, research and policy debates, with the author's own analysis of safety questions.
Worth knowing: Posts are long and assume some background knowledge.
Tool or datasetFor the curious
The Collective Intelligence Project · Weval · weval.org
Open platform where experts and communities write tests for AI models and publish the results, including checks on mental-health crisis responses and sycophancy.
Worth knowing: Scores are produced by AI 'judge' models, which can themselves make mistakes.