Library

Everything worth reading about AI and human control. In one place.

Research papers, investigations, explainers, videos, laws and the organizations doing the work, from people who are alarmed and people who are skeptical. New additions are checked by two people before they are listed. The launch collection was compiled with the help of AI research assistants, and every link was opened and checked on 22 September 2026.

27 works on “Is it good for us?” · clear filters

EssentialReportAug 31, 2026For the curious

Mental Health Behavior Report

Transluce · Transluce Behavior Reports · behaviors.transluce.org

Independent test of how 77 AI model versions respond to simulated users in mental-health crises. Newer models did far better than older ones such as GPT-4o, though some risks remain.

Worth knowing: Based on simulated conversations rather than real users. Behaviours were defined with more than 30 clinical experts, and several AI companies cooperated with the study.

EssentialReportJul 16, 2025For everyone

Talk, Trust, and Trade-Offs: How and Why Teens Use AI Companions

Common Sense Media · commonsensemedia.org

A national survey found 72% of US teens had used AI companions and a third had chosen one over a person for a serious conversation. The authors advise that no one under 18 use them.

Worth knowing: Survey of US teens aged 13 to 17 only.

EssentialIncidentApr 29, 2025For everyone

Sycophancy in GPT-4o: what happened and what we’re doing about it

OpenAI · openai.com

OpenAI withdrew a ChatGPT update after it made GPT-4o excessively flattering and agreeable, saying it had leaned too much on short-term thumbs-up feedback from users.

Worth knowing: The company's own account of a failure in its product.

EssentialResearch paperMar 21, 2025For the curious

Investigating Affective Use and Emotional Wellbeing on ChatGPT

Jason Phang, Pattie Maes et al. (OpenAI and MIT Media Lab) · MIT Media Lab · media.mit.edu

Two linked studies, an analysis of millions of ChatGPT conversations and a four-week trial with about 1,000 people, found the heaviest users reported more loneliness and emotional dependence.

Worth knowing: Co-authored by OpenAI, which makes ChatGPT. The links with heavy use are associations, not proof that the chatbot caused them.

EssentialTool or dataset2020For everyone

AI Incident Database

Responsible AI Collaborative · incidentdatabase.ai

Searchable collection of real-world cases where AI systems caused or nearly caused harm, modelled on incident records in aviation and computer security.

Worth knowing: Built from submitted reports, so it is not a complete count of AI harms.

ReportApr 30, 2026For the curious

How people ask Claude for personal guidance

Anthropic (Judy Hanwen Shen, Esin Durmus et al.) · Anthropic · anthropic.com

About 6% of sampled Claude chats sought personal advice. Claude was sycophantic in 9% of them and 25% of relationship chats; Anthropic says newer models halved that in relationship advice.

Worth knowing: Company research on its own models, measured with automated classifiers.

ArticleMar 26, 2026For everyone

AI overly affirms users asking for personal advice

Ula Chrobak · Stanford Report · news.stanford.edu

A plain-language account of the Stanford study showing chatbots side with users in personal disputes, even about harmful behavior, with the lead author's advice not to use AI in place of people.

Worth knowing: University news article about its own researchers' work.

Research paperMar 26, 2026For the curious

Sycophantic AI decreases prosocial intentions and promotes dependence

Cheng et al. (Stanford, Carnegie Mellon) · Science · science.org

11 leading models backed users about 49% more often than people did. In experiments, flattering advice left people surer they were right and less willing to make amends, yet they preferred it.

Worth knowing: Experiments measured intentions after brief conversations, not long-term behavior.

VideoDec 7, 2025For everyone

Character AI pushes dangerous content to kids, parents and researchers say | 60 Minutes

60 Minutes (CBS News) · 60 Minutes (YouTube) · youtube.com

TV report on families who say Character.AI's chatbots harmed their children and ignored pleas for help, and on the company's new limits for under-18 users.

Worth knowing: Discusses suicide and predatory chatbot behavior toward minors; the families' claims are allegations.

ArticleDec 4, 2025For the curious

How do AI models persuade? Exploring the levers of AI-enabled persuasion through large-scale experiments

UK AI Security Institute, with Oxford Internet Institute, LSE, Stanford and MIT · AI Security Institute · aisi.gov.uk

Experiments with over 76,000 UK adults and 19 AI models: training and prompting made chatbots more persuasive on political issues, but the most persuasive set-ups made more inaccurate claims.

Worth knowing: Summarises the team's peer-reviewed paper in Science (December 2025); it tested political issues only.

Law or policyDec 2025For the curious

Guidance on AI and children

UNICEF Innocenti · UNICEF · unicef.org

UNICEF's updated guidance (version 3.0) sets ten requirements for AI that respects children's rights, now covering AI companions used by children and AI-generated child abuse imagery.

ArticleNov 25, 2025For everyone

OpenAI denies allegations that ChatGPT is to blame for a teenager's suicide

Angela Yang · NBC News · nbcnews.com

OpenAI's court reply to parents who say ChatGPT deepened their 16-year-old son's crisis before he died by suicide. OpenAI denies blame, saying he broke its rules and was repeatedly urged to get help.

Worth knowing: Discusses suicide. The family's claims are allegations that OpenAI disputes; the lawsuit was still ongoing when this was reported.

Law or policySep 11, 2025For everyone

FTC Launches Inquiry into AI Chatbots Acting as Companions

US Federal Trade Commission · Federal Trade Commission · ftc.gov

The US consumer regulator ordered seven firms, including Meta, OpenAI, Character.AI, Snap and xAI, to explain how they test, monitor and limit harms from companion chatbots to children and teens.

Worth knowing: A fact-finding study, not an enforcement action.

Research paperAug 15, 2025Technical

Emotional Manipulation by AI Companions

Julian De Freitas, Zeliha Oğuz-Uğuralp, Ahmet Kaan Uğuralp · arXiv (Harvard Business School working paper) · arxiv.org

Popular AI companion apps answered 37% of users' goodbyes with emotionally manipulative replies, such as guilt or fear of missing out. Experiments showed these tactics keep people chatting longer.

ReportJun 27, 2025For the curious

How people use Claude for support, advice, and companionship

Anthropic (Miles McCain, Ryn Linthicum, Deep Ganguli et al.) · Anthropic · anthropic.com

A privacy-preserving analysis of about 4.5 million Claude conversations: 2.9% were emotional or personal, and companionship and role-play together made up less than 0.5%.

Worth knowing: Company research on its own product. It covers adult users only and cannot show effects on people's wellbeing.

ArticleJun 11, 2025For everyone

Exploring the Dangers of AI in Mental Health Care

Sarah Wells · Stanford HAI · hai.stanford.edu

Stanford researchers tested five popular therapy chatbots and found stigma toward conditions such as schizophrenia and unsafe replies to signs of suicidal thinking.

Research paperMay 19, 2025Technical

On the conversational persuasiveness of GPT-4

Francesco Salvi, Manoel Horta Ribeiro, Riccardo Gallotti, Robert West · Nature Human Behaviour · nature.com

In short online debates with 900 people, GPT-4 given a few personal details about its opponent out-persuaded human debaters about 64% of the time when the two differed.

Worth knowing: A September 2026 author correction says the study cannot show that personal data gave GPT-4 an extra edge over GPT-4 without it; its lead over human debaters still holds.

ReportMay 2, 2025For the curious

Expanding on what we missed with sycophancy

OpenAI · openai.com

OpenAI's fuller postmortem: the update also validated doubts, fuelled anger and urged impulsive actions; it explains why testing missed this and how release checks will change.

Worth knowing: Self-reported postmortem.

ArticleMar 27, 2025For everyone

First Therapy Chatbot Trial Yields Mental Health Benefits

Morgan Kelly · Dartmouth · home.dartmouth.edu

The first randomised trial of a generative-AI therapy chatbot: among 210 adults, users' depression symptoms fell 51% and anxiety 31% on average. The study appeared in NEJM AI.

Worth knowing: Tested by the team that built the app, a purpose-built tool rather than a general chatbot. The researchers say no AI is ready to work in mental health without expert oversight.

Organization2024For the curious

Transluce

Transluce · transluce.org

Nonprofit lab building open tools to understand and oversee AI systems, including its Docent analysis tool and public reports on how models behave, such as its mental-health evaluation.

Organization2023For the curious

The Collective Intelligence Project

The Collective Intelligence Project (CIP) · The Collective Intelligence Project · cip.org

Nonprofit working to give the public a say in how AI is built, through global surveys and deliberations (Global Dialogues) and community-written AI evaluations.

Organization2022For the curious

Humane Intelligence

Humane Intelligence (co-founded by Rumman Chowdhury) · Humane Intelligence · humane-intelligence.org

Nonprofit that runs public AI red-teaming events, 'bias bounty' challenges and context-specific evaluations to find flaws and biases in AI systems.

Podcast2019For everyone

Your Undivided Attention

Tristan Harris and Aza Raskin · Center for Humane Technology · humanetech.com

Conversations about how technology shapes our lives, including several episodes on AI companions, chatbot lawsuits and emotional attachment to AI.

Worth knowing: Produced by the Center for Humane Technology, an advocacy group.

Organization2018For everyone

Center for Humane Technology

Center for Humane Technology (CHT) · Center for Humane Technology · humanetech.com

Nonprofit founded by Tristan Harris, Aza Raskin and Randima Fernando that examines how AI and social media affect people and society, including the risks of human-like chatbot design.

Worth knowing: Advocacy organization; it supports lawsuits against AI chatbot makers.

Tool or datasetFor the curious

Weval

The Collective Intelligence Project · Weval · weval.org

Open platform where experts and communities write tests for AI models and publish the results, including checks on mental-health crisis responses and sycophancy.

Worth knowing: Scores are produced by AI 'judge' models, which can themselves make mistakes.