Back

Detection of Suicidality Through Privacy-Preserving Large Language Models

Wiest, I. C.; Verhees, F. G.; Ferber, D.; Zhu, J.; Bauer, M.; Lewitzka, U.; Pfennig, A.; Mikolas, P.; Kather, J. N.

2024-03-08 psychiatry and clinical psychology
10.1101/2024.03.06.24303763 medRxiv
Show abstract

ImportanceAttempts to use Artificial Intelligence (AI) in psychiatric disorders show moderate success, high-lighting the potential of incorporating information from clinical assessments to improve the models. The study focuses on using Large Language Models (LLMs) to manage unstructured medical text, particularly for suicide risk detection in psychiatric care. ObjectiveThe study aims to extract information about suicidality status from the admission notes of electronic health records (EHR) using privacy-sensitive, locally hosted LLMs, specifically evaluating the efficacy of Llama-2 models. Main Outcomes and MeasuresThe study compares the performance of several variants of the open source LLM Llama-2 in extracting suicidality status from psychiatric reports against a ground truth defined by human experts, assessing accuracy, sensitivity, specificity, and F1 score across different prompting strategies. ResultsA German fine-tuned Llama-2 model showed the highest accuracy (87.5%), sensitivity (83%) and specificity (91.8%) in identifying suicidality, with significant improvements in sensitivity and specificity across various prompt designs. Conclusions and RelevanceThe study demonstrates the capability of LLMs, particularly Llama-2, in accurately extracting the information on suicidality from psychiatric records while preserving data-privacy. This suggests their application in surveillance systems for psychiatric emergencies and improving the clinical management of suicidality by improving systematic quality control and research. Key PointsO_ST_ABSQuestionC_ST_ABSCan large language models (LLMs) accurately extract information on suicidality from electronic health records (EHR)? FindingsIn this analysis of 100 psychiatric admission notes using Llama-2 models, the German fine-tuned model (Emgerman) demonstrated the highest accuracy (87.5%), sensitivity (83%) and specificity (91.8%) in identifying suicidality, indicating the models effectiveness in on-site processing of clinical documentation for suicide risk detection. MeaningThe study highlights the effectiveness of LLMs, particularly Llama-2, in accurately extracting the information on suicidality from psychiatric records, while preserving data privacy. It recommends further evaluating these models to integrate them into clinical management systems to improve detection of psychiatric emergencies and enhance systematic quality control and research in mental health care.

Matching journals

The top 9 journals account for 50% of the predicted probability mass.

1
npj Digital Medicine
118 papers in training set
Top 0.6%
11.6%
2
Acta Psychiatrica Scandinavica
10 papers in training set
Top 0.1%
7.7%
3
Frontiers in Psychiatry
87 papers in training set
Top 0.2%
7.7%
4
Psychiatry Research
41 papers in training set
Top 0.1%
6.5%
5
Frontiers in Digital Health
24 papers in training set
Top 0.2%
5.3%
6
Journal of Medical Internet Research
87 papers in training set
Top 0.7%
3.9%
7
The British Journal of Psychiatry
23 papers in training set
Top 0.1%
3.9%
8
Acta Neuropsychiatrica
14 papers in training set
Top 0.1%
3.1%
9
Translational Psychiatry
260 papers in training set
Top 2%
3.1%
50% of probability mass above
10
Schizophrenia
21 papers in training set
Top 0.2%
2.6%
11
Journal of Affective Disorders
92 papers in training set
Top 0.9%
2.3%
12
PLOS ONE
5266 papers in training set
Top 44%
2.3%
13
Scientific Reports
3612 papers in training set
Top 45%
2.3%
14
JMIR Medical Informatics
18 papers in training set
Top 0.4%
2.1%
15
JAMIA Open
42 papers in training set
Top 0.8%
1.9%
16
JAMA Psychiatry
15 papers in training set
Top 0.2%
1.9%
17
BMC Psychiatry
25 papers in training set
Top 0.4%
1.9%
18
European Psychiatry
11 papers in training set
Top 0.1%
1.7%
19
BJPsych Open
29 papers in training set
Top 0.4%
1.5%
20
JMIR Formative Research
33 papers in training set
Top 0.9%
1.4%
21
Nature Medicine
125 papers in training set
Top 2%
1.3%
22
Frontiers in Artificial Intelligence
20 papers in training set
Top 0.5%
1.1%
23
Psychological Medicine
88 papers in training set
Top 2%
1.1%
24
Epidemiology and Psychiatric Sciences
11 papers in training set
Top 0.3%
1.0%
25
Journal of the American Medical Informatics Association
71 papers in training set
Top 2%
1.0%
26
Schizophrenia Bulletin
32 papers in training set
Top 0.4%
1.0%
27
BMC Medical Informatics and Decision Making
43 papers in training set
Top 2%
1.0%
28
BMJ Mental Health
15 papers in training set
Top 0.4%
1.0%
29
Communications Medicine
113 papers in training set
Top 4%
1.0%
30
Computational Psychiatry
12 papers in training set
Top 0.2%
0.8%