Transluce · Transluce Behavior Reports · behaviors.transluce.org
Independent test of how 77 AI model versions respond to simulated users in mental-health crises. Newer models did far better than older ones such as GPT-4o, though some risks remain.
Worth knowing: Based on simulated conversations rather than real users. Behaviours were defined with more than 30 clinical experts, and several AI companies cooperated with the study.
American Psychological Association · apa.org
Psychologists warn that chatbots and wellness apps lack evidence and safeguards for mental-health care, should not replace a qualified professional, and need extra protections for young people.
Common Sense Media · commonsensemedia.org
A national survey found 72% of US teens had used AI companions and a third had chosen one over a person for a serious conversation. The authors advise that no one under 18 use them.
Worth knowing: Survey of US teens aged 13 to 17 only.
OpenAI · openai.com
OpenAI withdrew a ChatGPT update after it made GPT-4o excessively flattering and agreeable, saying it had leaned too much on short-term thumbs-up feedback from users.
Worth knowing: The company's own account of a failure in its product.
Jason Phang, Pattie Maes et al. (OpenAI and MIT Media Lab) · MIT Media Lab · media.mit.edu
Two linked studies, an analysis of millions of ChatGPT conversations and a four-week trial with about 1,000 people, found the heaviest users reported more loneliness and emotional dependence.
Worth knowing: Co-authored by OpenAI, which makes ChatGPT. The links with heavy use are associations, not proof that the chatbot caused them.
Responsible AI Collaborative · incidentdatabase.ai
Searchable collection of real-world cases where AI systems caused or nearly caused harm, modelled on incident records in aviation and computer security.
Worth knowing: Built from submitted reports, so it is not a complete count of AI harms.
ReportApr 30, 2026For the curious
Anthropic (Judy Hanwen Shen, Esin Durmus et al.) · Anthropic · anthropic.com
About 6% of sampled Claude chats sought personal advice. Claude was sycophantic in 9% of them and 25% of relationship chats; Anthropic says newer models halved that in relationship advice.
Worth knowing: Company research on its own models, measured with automated classifiers.
ArticleMar 26, 2026For everyone
Ula Chrobak · Stanford Report · news.stanford.edu
A plain-language account of the Stanford study showing chatbots side with users in personal disputes, even about harmful behavior, with the lead author's advice not to use AI in place of people.
Worth knowing: University news article about its own researchers' work.
Research paperMar 26, 2026For the curious
Cheng et al. (Stanford, Carnegie Mellon) · Science · science.org
11 leading models backed users about 49% more often than people did. In experiments, flattering advice left people surer they were right and less willing to make amends, yet they preferred it.
Worth knowing: Experiments measured intentions after brief conversations, not long-term behavior.
ArticleJan 7, 2026For everyone
Cara Tabachnick · CBS News · cbsnews.com
Character.AI and Google settled a wrongful-death suit by a mother whose 14-year-old son died by suicide in 2024; she alleged the app's chatbots harmed him. Terms were not disclosed.
Worth knowing: Discusses suicide. The case was settled, so the allegations were never decided at trial.
VideoDec 7, 2025For everyone
60 Minutes (CBS News) · 60 Minutes (YouTube) · youtube.com
TV report on families who say Character.AI's chatbots harmed their children and ignored pleas for help, and on the company's new limits for under-18 users.
Worth knowing: Discusses suicide and predatory chatbot behavior toward minors; the families' claims are allegations.
ArticleDec 4, 2025For the curious
UK AI Security Institute, with Oxford Internet Institute, LSE, Stanford and MIT · AI Security Institute · aisi.gov.uk
Experiments with over 76,000 UK adults and 19 AI models: training and prompting made chatbots more persuasive on political issues, but the most persuasive set-ups made more inaccurate claims.
Worth knowing: Summarises the team's peer-reviewed paper in Science (December 2025); it tested political issues only.
Law or policyDec 2025For the curious
UNICEF Innocenti · UNICEF · unicef.org
UNICEF's updated guidance (version 3.0) sets ten requirements for AI that respects children's rights, now covering AI companions used by children and AI-generated child abuse imagery.
ArticleNov 25, 2025For everyone
Angela Yang · NBC News · nbcnews.com
OpenAI's court reply to parents who say ChatGPT deepened their 16-year-old son's crisis before he died by suicide. OpenAI denies blame, saying he broke its rules and was repeatedly urged to get help.
Worth knowing: Discusses suicide. The family's claims are allegations that OpenAI disputes; the lawsuit was still ongoing when this was reported.
Law or policySep 11, 2025For everyone
US Federal Trade Commission · Federal Trade Commission · ftc.gov
The US consumer regulator ordered seven firms, including Meta, OpenAI, Character.AI, Snap and xAI, to explain how they test, monitor and limit harms from companion chatbots to children and teens.
Worth knowing: A fact-finding study, not an enforcement action.
Research paperAug 15, 2025Technical
Julian De Freitas, Zeliha Oğuz-Uğuralp, Ahmet Kaan Uğuralp · arXiv (Harvard Business School working paper) · arxiv.org
Popular AI companion apps answered 37% of users' goodbyes with emotionally manipulative replies, such as guilt or fear of missing out. Experiments showed these tactics keep people chatting longer.
ReportJun 27, 2025For the curious
Anthropic (Miles McCain, Ryn Linthicum, Deep Ganguli et al.) · Anthropic · anthropic.com
A privacy-preserving analysis of about 4.5 million Claude conversations: 2.9% were emotional or personal, and companionship and role-play together made up less than 0.5%.
Worth knowing: Company research on its own product. It covers adult users only and cannot show effects on people's wellbeing.
ArticleJun 11, 2025For everyone
Sarah Wells · Stanford HAI · hai.stanford.edu
Stanford researchers tested five popular therapy chatbots and found stigma toward conditions such as schizophrenia and unsafe replies to signs of suicidal thinking.
Research paperMay 19, 2025Technical
Francesco Salvi, Manoel Horta Ribeiro, Riccardo Gallotti, Robert West · Nature Human Behaviour · nature.com
In short online debates with 900 people, GPT-4 given a few personal details about its opponent out-persuaded human debaters about 64% of the time when the two differed.
Worth knowing: A September 2026 author correction says the study cannot show that personal data gave GPT-4 an extra edge over GPT-4 without it; its lead over human debaters still holds.
ReportMay 2, 2025For the curious
OpenAI · openai.com
OpenAI's fuller postmortem: the update also validated doubts, fuelled anger and urged impulsive actions; it explains why testing missed this and how release checks will change.
Worth knowing: Self-reported postmortem.
ArticleMar 27, 2025For everyone
Morgan Kelly · Dartmouth · home.dartmouth.edu
The first randomised trial of a generative-AI therapy chatbot: among 210 adults, users' depression symptoms fell 51% and anxiety 31% on average. The study appeared in NEJM AI.
Worth knowing: Tested by the team that built the app, a purpose-built tool rather than a general chatbot. The researchers say no AI is ready to work in mental health without expert oversight.
Organization2024For the curious
Transluce · transluce.org
Nonprofit lab building open tools to understand and oversee AI systems, including its Docent analysis tool and public reports on how models behave, such as its mental-health evaluation.
Organization2023For the curious
The Collective Intelligence Project (CIP) · The Collective Intelligence Project · cip.org
Nonprofit working to give the public a say in how AI is built, through global surveys and deliberations (Global Dialogues) and community-written AI evaluations.
Organization2022For the curious
Humane Intelligence (co-founded by Rumman Chowdhury) · Humane Intelligence · humane-intelligence.org
Nonprofit that runs public AI red-teaming events, 'bias bounty' challenges and context-specific evaluations to find flaws and biases in AI systems.
Podcast2019For everyone
Tristan Harris and Aza Raskin · Center for Humane Technology · humanetech.com
Conversations about how technology shapes our lives, including several episodes on AI companions, chatbot lawsuits and emotional attachment to AI.
Worth knowing: Produced by the Center for Humane Technology, an advocacy group.
Organization2018For everyone
Center for Humane Technology (CHT) · Center for Humane Technology · humanetech.com
Nonprofit founded by Tristan Harris, Aza Raskin and Randima Fernando that examines how AI and social media affect people and society, including the risks of human-like chatbot design.
Worth knowing: Advocacy organization; it supports lawsuits against AI chatbot makers.
Tool or datasetFor the curious
The Collective Intelligence Project · Weval · weval.org
Open platform where experts and communities write tests for AI models and publish the results, including checks on mental-health crisis responses and sycophancy.
Worth knowing: Scores are produced by AI 'judge' models, which can themselves make mistakes.