Research Letter: Application of GPT-4 to select next-step antidepressant treatment in major depression
Perlis, R. H.
Show abstract
IntroductionLarge language models perform well on a range of academic tasks including medical examinations. The performance of this class of models in psychopharmacology has not been explored. MethodChat GPT-plus, implementing the GPT-4 large language model, was presented with each of 10 previously-studied antidepressant prescribing vignettes in randomized order, with results regenerated 5 times to evaluate stability of responses. Results were compared to expert consensus. ResultsAt least one of the optimal medication choices was included among the best choices in 38/50 (76%) vignettes: 5/5 for 7 vignettes, 3/5 for 1, and 0/5 for 2. At least one of the poor choice or contraindicated medications was included among the choices considered optimal or good in 24/50 (48%) of vignettes. The model provided as rationale for treatment selection multiple heuristics including avoiding prior unsuccessful medications, avoiding adverse effects based on comorbidities, and generalizing within medication class. ConclusionThe model appeared to identify and apply a number of heuristics commonly applied in psychopharmacologic clinical practice. However, the inclusion of less optimal recommendations indicates that large language models may pose a substantial risk if routinely applied to guide psychopharmacologic treatment without further monitoring.
Matching journals
The top 11 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Sociodemographic, clinical, and genetic factors associated with self-reported antidepressant response outcomes in the UK Biobank 92%
- Antipsychotic Polypharmacy and Adverse Drug Reactions Among Adults in a London Mental Health Service, 2008-2018 92%
- Predicting involuntary admission following inpatient psychiatric treatment using machine learning trained on electronic health record data 92%
Similar papers in this journal
- UK medical students’ self-reported knowledge and harm assessment of psychedelics and their application in clinical research: a cross-sectional study 93%
- The Antidepressant Advisor (ADeSS): A Decision Support System for Antidepressant Treatment for Depression in UK primary care – a feasibility study 93%
- Global prevalence of antidepressant utilization in the community: A protocol for a systematic review 93%
Similar papers in this journal
- Applications of Large Language Models in Psychiatry: A Systematic Review 93%
- Development of Goal Management Training + (GMT + ) for Methamphetamine Use Disorder Through Collaborative Design: A Process Description 91%
- Patients with affective disorders profit most from telemedical treatment: Evidence from a naturalistic patient cohort during the COVID-19 pandemic 90%
Similar papers in this journal
- Predicting remission after internet-delivered psychotherapy in patients with depression using machine learning and multi-modal data 93%
- Pharmacologic and genetic evidence converge on mechanisms of psychotic illness 92%
- Identifying Psychosis Episodes in Psychiatric Admission Notes via Rule-based Methods, Machine Learning, and Pre-Trained Language Models 91%
Similar papers in this journal
- Unveiling the Hidden Toll of Drug-Induced Impulsivity: A Network Analysis of the FDA Adverse Event Reporting System 92%
- Patient-Reported Reasons for Antihypertensive Medication Change: A Quantitative Study Using Social Media 91%
- Standardization of drug names in the FDA Adverse Event Reporting System: The DiAna dictionary 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.