Characterizing artificial intelligence (AI) psychosis in a large academic medical setting: evidence of the new clinical phenomenon and the vulnerability of those in early phases of psychosis
Bergson, Z.; Vassall, S. G.; Wright, A.; McCoy, A. B.; Schafer, K. M.; Achee, M. C.; Sheffield, J. M.
Show abstract
Background: Concerns about "AI psychosis" have swirled in the media since ChatGPT's release, but few systematic analyses exist. We therefore conducted an electronic health record (EHR) analysis to identify the frequency, clinical characteristics, and quality of AI interactions in patients experiencing psychosis treated in a medical center. Methods: AI keywords (e.g., ChatGPT, AI) were used to search Vanderbilt University Medical Center's EHR from 12/1/2022-4/1/2026. Records were discarded if they were not AI-related or if the primary diagnosis did not include psychosis. Three raters read notes to determine if a patient was experiencing AI psychosis and classified the interactions using 4 a-priori categories (Catalyst, Amplifier, Co-Author, Object) formulated to explain how AI-related negative outcomes emerge. Findings: 73 patients met our criteria. 28 patients were rated as experiencing AI psychosis, 17 had neutral interactions, and 28 expressed delusional content related to AI without documented evidence of conversational AI use. ChatGPT was the matching keyword for 53.6% patients experiencing AI psychosis. The majority of AI psychosis cases were documented after ChatGPT's "4o" model was released in May 2024. Notably, the AI Psychosis group had significantly more patients experiencing a first psychotic episode (60.7%) compared to the other two groups. Amplifier was the most common (64.3%) qualitative rating in the AI Psychosis group. Interpretation: "AI psychosis" is an infrequent but real phenomenon observed in clinical practice. Most affected patients were experiencing their first psychotic episode and presented with AI psychosis following the release of the more sycophantic GPT-4o. Among the affected patients, AI most often exacerbated an existing condition by reinforcing distorted ideas.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Systematic Review on Psychological and Biological Mediators Between Adversity and Psychosis: Potential Targets for Treatment 92%
- Jumping To Conclusions, General Intelligence, And Psychosis Liability: Findings From The Multicentric EU-GEI Case-Control Study 92%
- Predicting involuntary admission following inpatient psychiatric treatment using machine learning trained on electronic health record data 91%
Similar papers in this journal
- Analysis of diagnosis instability in electronic health records reveals diverse disease trajectories of severe mental illness 94%
- Associations between antipsychotic use, substance use and relapse risks in patients with schizophrenia - real-world evidence from two national cohorts 93%
- Latent subtypes of manic or irritable episode symptoms in two population-based cohorts 92%
Similar papers in this journal
- Evidence for feasibility of mobile health and social media-based interventions for early psychosis and clinical high risk 94%
- Facial and vocal markers of schizophrenia measured using remote smartphone assessments 90%
- Development of the NeuroFlow Severity Score and Comparison With Validated Measures for Depression and Anxiety 90%
Similar papers in this journal
- Validation of an ICD-code-based case definition for psychotic illness across three health systems 97%
- Can we detect the undetected? Comparing the prodromes of individuals with first episode psychosis detected and undetected by clinical high risk for psychosis services: an electronic health record study 96%
- Latent Factors of Language Disturbance and Relationships to Quantitative Speech Features 94%
Similar papers in this journal
- Short-term functional outcome in psychotic patients. Results of the Turku Early Psychosis Study (TEPS) 92%
- Current state of the evidence on community treatments for people with complex emotional needs: a scoping review 91%
- Treatment Resistant Depression in electronic health records: Definitions Matter 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.