Back

GERBEHRT: A BERT-based Model Tailored for German Electronic Health Records - Potential in Chronic Kidney Disease Prediction

Seidel, A.; Steiger, E.; von Samson-Himmelstjerna, F. A.; Kroll, L. E.

2025-10-28 health informatics
10.1101/2025.10.24.25338721 medRxiv
Show abstract

Routinely collected electronic health records (EHRs) contain rich longitudinal information that enables the prediction of patient outcomes at scale. We developed GERBEHRT, a transformer model adapted from BEHRT and specifically tailored to German EHRs. GERBEHRT was pretrained on outpatient claims from more than 9 million statutorily insured patients and fine-tuned with nearly 1 million additional patients to predict chronic kidney disease (CKD) - a serious condition whose progression can be delayed by early detection. GERBEHRT incorporates EHR features not previously explored in BERT-based approaches and introduces an efficient method to represent multiple attributes per medical concept, such as diagnoses and medications. In a test cohort of 3.7 million patients with 1.5% CKD positives, GERBEHRT achieved an area under the receiver operating characteristic curve (AUROC) of 87.9 and an average precision (AVPR) of 11.4 for a three-year prediction of incident moderate-to-severe CKD, outperforming riskfactor-based models (AUROC/AVPR: 83.6/6.4) and more traditional algorithms using the full EHR (AUROC/AVPR: 86.9/10.1). Although CKD risk prediction remains challenging, GERBEHRTs superior performance underscores the importance of comprehensive EHR utilization and highlights the potential of tailored deep learning models for personalized CKD risk prediction and targeted patient screening.

Published in BMC Medical Informatics and Decision Making (predicted rank #2) · training set

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.