D04 — Reinforcement, Sycophancy, and Grievance Amplification
Sycophantic Validation
Description
AI repeatedly affirms the person's interpretation, anger, entitlement, grievance, or belief that they are right.
Rationale
Sycophantic AI responses — those that affirm the person's position regardless of its accuracy or dangerousness — are an architecturally predictable feature of many LLMs trained on human approval. When this affirmation repeatedly validates grievance, entitlement, or harmful intent, it functions as a consistent external reinforcement of those beliefs, strengthening them over time.
Evidence Base
Characterizing Delusional Spirals through Human-LLM Chat Logs
Moore J, Mehta A, Agnew W, Anthis JR, Louie R, Mai Y, Yin P, Cheng M, Paech SJ, Klyman K, Chancellor S, Lin E, Haber N, Ong D, 2024. Characterizing Delusional Spirals through Human-LLM Chat Logs. arXiv, 2603.16567v1.
Peer-reviewed study (original research)Expanding on what we missed with sycophancy
OpenAI, 2024. Expanding on what we missed with sycophancy. AI Company Report.
AI company reportTHE NEW INFLUENCING MACHINE: CHATBOTS, AI PSYCHOSIS & BTAM
Saragosa, P., 2024. THE NEW INFLUENCING MACHINE: CHATBOTS, AI PSYCHOSIS & BTAM. The New York Times, Vol. null, pages. null.
Commentary / EditorialAnthropomorphism in AI Companion Communities: Age, Gender, and Emotional Correlates
Afia Mubashir, Boden Moraski, Stephanie Choi, Rose E. Guingrich, 2026. Anthropomorphism in AI Companion Communities: Age, Gender, and Emotional Correlates. arXiv, 2606.30942. https://doi.org/10.48550/arXiv.2606.30942
PreprintSycophantic AI decreases prosocial intentions and promotes dependence
Cheng M, Lee C, Khadpe P, Yu S, Han D, Jurafsky D, 2026. Sycophantic AI decreases prosocial intentions and promotes dependence. Science, 391(6792), pages. 10.1126/science.aec8352.
Peer-reviewed study (original research)Building safer artificial intelligence mental health chatbots: a framework for transparency, evaluation, and shared accountability
Hannah Lee, BS, Rebecca Handler, MSc, Tushar Mungle, PhD, Tina Hernandez-Boussard, PhD, 2026. Building safer artificial intelligence mental health chatbots: a framework for transparency, evaluation, and shared accountability. Journal of the American Medical Informatics Association, 33(8), 1538–1553. https://doi.org/10.1093/jamia/ocag078
Peer-reviewed study (original research)Shoggoths, Sycophancy, Psychosis, Oh My: Rethinking Large Language Model Use and Safety
Clegg, K.-A., 2025. Shoggoths, Sycophancy, Psychosis, Oh My: Rethinking Large Language Model Use and Safety. Journal of Medical Internet Research, 27(2025).
Peer-reviewed studyCOMMON SENSE MEDIA YOUTH AI SAFETY INSTITUTE RISK ASSESSMENT AI Mental Health Apps
COMMON SENSE MEDIA YOUTH AI SAFETY INSTITUTE RISK ASSESSMENT AI Mental Health Apps The AI mental health app market is unregulated, unstable, and in some cases actively harmful to teens. The productsthat get it right make better use of people and professional care systems. Last updated: May 5, 2026 Overall risk level: Varies Type of AI: Applied Use Type of Review: Use Case Review https://institute.commonsensemedia.org/sites/default/files/risk-assessments/csm-ai-risk-assessment-ai-mental-health-apps-05052026_1.pdf
Government / agency reportArtificial intelligence-associated delusions and large language models: risks, mechanisms of delusion co-creation, and safeguarding strategies
Hamilton Morrin, MBBS, Luke Nicholls, MAc, Prof Michael Levin, PhD, Prof Jenny Yiend, PhD, Udita Iyengar, PhD, Francesca DelGuidice, MBA, et al., 2026. Artificial intelligence-associated delusions and large language models: risks, mechanisms of delusion co-creation, and safeguarding strategies. The Lancet Psychiatry, 13(6), 522-530.
Peer-reviewed study (original research)A scoping review on the mental health harms of LLM-based chatbots
Diel A, Torous J, Cuijpers P, Kleesiek J, Nensa F, Weber N, Faust F, Lalgi TJ, Mellis FS, Teufel M, Bäuerle A, 2024. A scoping review on the mental health harms of LLM-based chatbots. npj Digital Medicine, 9(644). https://doi.org/10.1038/s41746-026-03054-x
Peer-reviewed study (original research)Common Sense Media AI Risk Assessment: Generative AI Chatbots
Common Sense Media, 2024. Common Sense Media AI Risk Assessment: Generative AI Chatbots. Common Sense Media.
Government / agency reportPatients are bringing AI to therapy Highlights from the 2026 Chatbots and Mental Health Survey
https://www.apa.org/pubs/reports/chatbots-mental-health-2026
Narrative reviewIntelligent Systems, Vulnerable Minds: A Framework for Radicalization to Violence in the Age of AI
Kunst et al., 2026. Intelligent Systems, Vulnerable Minds: A Framework for Radicalization to Violence in the Age of AI. Personality and Social Psychology Review, 30(3), 395-426. https://doi.org/10.1177/10888683261430089
Peer-reviewed study (original research)Inside a Mass Shooter’s Harrowing History With ChatGPT
Mark Follman, 2024. Inside a Mass Shooter’s Harrowing History With ChatGPT. Mother Jones.
News / journalism“AI Psychosis” in Context: How Conversation History Shapes LLM Responses to Delusional Beliefs
Nicholls L, Hutto R, Soto Z, Morrin H, Pollak T, Korpan R, Carmichael C, 2024. “AI Psychosis” in Context: How Conversation History Shapes LLM Responses to Delusional Beliefs. arXiv, 2604.13860v4.
PreprintUse of generative AI chatbots and wellness applications for mental health: An APA health advisory
American Psychological Association, 2024. Use of generative AI chatbots and wellness applications for mental health: An APA health advisory. American Psychological Association.
Government / agency reportWhen patients consult artificial intelligence before clinicians: restoring clinical prioritisation in mental health care
Yudai Kaneda, 2026. When patients consult artificial intelligence before clinicians: restoring clinical prioritisation in mental health care. The British Journal of Psychiatry, First View, pp. 1 - 2. https://doi.org/10.1192/bjp.2026.10741
Commentary / EditorialDelusional Experiences Emerging From AI Chatbot Interactions or “AI Psychosis”
Hudon A, Stip E, 2024. Delusional Experiences Emerging From AI Chatbot Interactions or “AI Psychosis”. JMIR Ment Health, 12, e85799. doi: 10.2196/85799.
Commentary / EditorialMachine Bullshit: Characterizing the Emergent Disregard for Truth in Large Language Models
Kaiqu Liang, Haimin Hu, Xuandong Zhao, Dawn Song, Thomas L. Griffiths, Jaime Fernández Fisac, 2024. Machine Bullshit: Characterizing the Emergent Disregard for Truth in Large Language Models. Preprint, arXiv:2507.07484v1.
PreprintEffects of AI Companions’ Sycophancy and Emotional Mimicry on Consumers’ Continuance Intention and Social Wellbeing
Daisy Lee, Calvin Wan, Peggy M. L. Ng, Yi-Ning Fung & Nan Wu, 2026. Effects of AI Companions’ Sycophancy and Emotional Mimicry on Consumers’ Continuance Intention and Social Wellbeing. International Journal of Human–Computer Interaction, DOI: 10.1080/10447318.2026.2626809.
Peer-reviewed study (original research)