Back

Optimized BERT-based NLP outperforms Zero-Shot Methods for Automated Symptom Detection in Clinical Practice

Diaz Ochoa, J. G.; Layer, N.; Mahr, J.; Mustafa, F. E.; Menzel, C. U.; Mueller-Schilling, M.; Schilling, T.; Illerhaus, G.; Knott, M.; Krohn, A.

2025-04-22 health informatics
10.1101/2025.04.21.25326037 medRxiv
Show abstract

AO_SCPLOWBSTRACTC_SCPLOWO_ST_ABSBO_SCPLOWACKGROUNDC_SCPLOWC_ST_ABSLarge Language Nodels (LLMs) have raised broad expectations for clinical use, particularly in the processing of complex medical narratives. However, in practice, more targeted Natural Language Processing (NLP) approaches may offer higher precision and feasibility for symptom extraction from real-world clinical texts. NLP provides promising tools for extracting clinical information from unstructured medical narratives. However, few studies have focused on integrating symptom information from free texts in German, particularly for complex patient groups such as emergency department (ED) patients. The ED setting presents specific challenges: high documentation pressure, heterogeneous language styles, and the need for secure, locally deployable models due to strict data protection regulations. Furthermore, German remains a low-resource language in clinical NLP. MO_SCPLOWETHODSC_SCPLOWWe implemented and compared two models for zero-shot learning--GLiNER and Mistral--and a fine-tuned BERT-based SCAI-BIO/BioGottBERT model for named entity recognition (NER) of symptoms, anatomical terms, and negations in German ED anamnesis texts in an on-premises environment in a hospital. Manual annotations of 150 narratives were used for model validation. The postprocessing steps included confidence-based filtering, negation exclusion, symptom standardization, and integration with structured oncology registry data. All computations were performed on local hospital servers in an on-premises implementation to ensure full data protection compliance. RO_SCPLOWESULTSC_SCPLOWThe fine-tuned SCAI-BIO/BioGottBERT model outperformed both zero-shot approaches, achieving an F1 score of 0.84 for symptom extraction and demonstrating superior performance in negation detection. The validated pipeline enabled systematic extraction of affirmed symptoms from ED-free text, transforming them into structured data. This method allows large-scale analysis of symptom profiles across patient populations and serves as a technical foundation for symptom-based clustering and subgroup analysis. CO_SCPLOWONCLUSIONSC_SCPLOWOur study demonstrates that modern NLP methods can reliably extract clinical symptoms from German ED free text, even under strict data protection constraints and with limited training resources. Fine-tuned models offer a precise and practical solution for integrating unstructured narratives into clinical decision-making. This work lays the methodological foundation for a new way of systematically analyzing large patient cohorts on the basis of free-text data. Beyond symptoms, this approach can be extended to extracting diagnoses, procedures, or other clinically relevant entities. Building upon this framework, we apply network-based clustering methods (in a subsequent study) to identify clinically meaningful patient subgroups and explore sex- and age-specific patterns in symptom expression.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Artificial Intelligence in Medicine
17 papers in training set
Top 0.1%
12.8%
2
Journal of Biomedical Informatics
47 papers in training set
Top 0.1%
11.7%
3
BMC Medical Informatics and Decision Making
43 papers in training set
Top 0.2%
7.8%
4
Journal of the American Medical Informatics Association
71 papers in training set
Top 0.5%
7.2%
5
npj Digital Medicine
118 papers in training set
Top 0.9%
6.2%
6
Bioinformatics
1204 papers in training set
Top 4%
6.2%
50% of probability mass above
7
JAMIA Open
42 papers in training set
Top 0.3%
5.4%
8
JMIR Medical Informatics
18 papers in training set
Top 0.1%
4.3%
9
BMJ Health & Care Informatics
15 papers in training set
Top 0.3%
3.2%
10
Scientific Reports
3612 papers in training set
Top 35%
3.2%
11
International Journal of Medical Informatics
26 papers in training set
Top 0.5%
2.6%
12
JCO Clinical Cancer Informatics
22 papers in training set
Top 0.3%
2.4%
13
Journal of Medical Internet Research
87 papers in training set
Top 1%
2.4%
14
Computer Methods and Programs in Biomedicine
28 papers in training set
Top 0.4%
2.1%
15
Computers in Biology and Medicine
128 papers in training set
Top 2%
2.1%
16
Patterns
78 papers in training set
Top 1%
2.1%
17
PLOS Digital Health
106 papers in training set
Top 3%
1.4%
18
Frontiers in Digital Health
24 papers in training set
Top 0.9%
1.4%
19
Biology Methods and Protocols
61 papers in training set
Top 1%
1.3%
20
PLOS ONE
5266 papers in training set
Top 57%
1.0%
21
iScience
1154 papers in training set
Top 30%
1.0%
22
Frontiers in Artificial Intelligence
20 papers in training set
Top 0.6%
1.0%
23
Scientific Data
209 papers in training set
Top 3%
0.8%
24
Bioinformatics Advances
203 papers in training set
Top 5%
0.6%
25
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 2%
0.6%
26
BMC Bioinformatics
457 papers in training set
Top 6%
0.6%