Back

Fine-Tuned Large Language Models for Detecting Social Isolation from Unstructured Clinical Notes

Chinthala, L. K.; Lemon, C.; Shaban-Nejad, A.; Farage, G.; Davis, R. L.; Xu, H.; Madlock-Brown, C.

2026-07-07 health informatics
10.64898/2026.07.05.26357334 medRxiv
Show abstract

Objectives: This study aimed to leverage FLAN-T5-Large, BERT, RoBERTa, and Gemma-2-2B, with fine-tuning, to identify instances of social isolation and social support within unstructured clinical notes. Materials and Methods: Annotated clinical note spans containing social context cues were used to fine-tune each model. Performance was evaluated using Accuracy, Precision, Recall, and Macro-F1 score. A structured prompt was used to instruct the model to perform classification task and mitigate overgeneralization. Performance comparisons across the models assessed sensitivity, robustness, and false positive reduction. Results: FLAN-T5-Large achieved highest performance, with Macro-F1 of 0.92{+/-}0.04, demonstrating balanced results across classes: social isolation (F1 = 0.91{+/-}0.03), no social isolation (F1 = 0.94{+/-}0.05), and social support (F1 = 0.90{+/-}0.04). Gemma-2-2B produced comparable results, with Macro-F1 score of 0.89{+/-}0.10. BERT and RoBERTa achieved lower Macro-F1 scores of 0.77{+/-}0.17 and 0.80{+/-}0.21 respectively, with variability across categories. Discussion: A major contribution of this work is precise identification of multiple concepts related to social connectedness. By integrating annotated examples of both true and false positives, including negations and contextually ambiguous terms, the model better distinguished relevant social context cues from noise. Training on both social isolation and support provided a dual framework for comparative analyses and patient stratification. Conclusion: Transformer-based NLP models, particularly FLAN-T5-Large, demonstrated potential for identifying social isolation and social support in clinical text. These findings support the use of generative AI techniques to enhance detection of social isolation from EHRs, advancing context-aware healthcare analytics.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
Journal of Biomedical Informatics
47 papers in training set
Top 0.1%
18.0%
2
Journal of the American Medical Informatics Association
71 papers in training set
Top 0.3%
11.6%
3
Frontiers in Digital Health
24 papers in training set
Top 0.1%
9.5%
4
npj Digital Medicine
118 papers in training set
Top 0.7%
9.4%
5
PLOS Digital Health
106 papers in training set
Top 0.7%
7.7%
50% of probability mass above
6
JAMIA Open
42 papers in training set
Top 0.2%
6.5%
7
BMC Medical Informatics and Decision Making
43 papers in training set
Top 0.3%
6.1%
8
JMIR Medical Informatics
18 papers in training set
Top 0.1%
4.2%
9
Artificial Intelligence in Medicine
17 papers in training set
Top 0.2%
3.2%
10
International Journal of Medical Informatics
26 papers in training set
Top 0.4%
2.7%
11
Journal of Medical Internet Research
87 papers in training set
Top 1%
2.3%
12
Scientific Reports
3612 papers in training set
Top 56%
1.7%
13
BMJ Health & Care Informatics
15 papers in training set
Top 0.6%
1.6%
14
Communications Medicine
113 papers in training set
Top 3%
1.3%
15
eBioMedicine
183 papers in training set
Top 4%
1.1%
16
Computers in Biology and Medicine
128 papers in training set
Top 3%
1.1%
17
DIGITAL HEALTH
17 papers in training set
Top 0.7%
1.1%
18
PLOS ONE
5266 papers in training set
Top 56%
1.1%
19
Computer Methods and Programs in Biomedicine
28 papers in training set
Top 0.9%
1.0%
20
Biology Methods and Protocols
61 papers in training set
Top 2%
1.0%
21
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 1%
1.0%
22
Patterns
78 papers in training set
Top 3%
0.8%
23
BMC Medical Research Methodology
47 papers in training set
Top 2%
0.6%
24
iScience
1154 papers in training set
Top 42%
0.6%