Exclusive: New Research Shows AI Strategically Lying
Billy Perrigo · TIME · time.com
An accessible report on the alignment-faking study, explaining why a model that misleads its trainers could make safety training harder to trust.
Research papers, investigations, explainers, videos, laws and the organizations doing the work, from people who are alarmed and people who are skeptical. New additions are checked by two people before they are listed. The launch collection was compiled with the help of AI research assistants, and every link was opened and checked on 22 September 2026.
1 work on “Does it tell the truth?” · Article · clear filters
An accessible report on the alignment-faking study, explaining why a model that misleads its trainers could make safety training harder to trust.