Back

Episode-Driven Insights: Can Large Language Models Tackle Multimodal Diabetes Data?

Choi, H. J.; Raj, S.

2025-04-25 endocrinology
10.1101/2025.04.24.25326385 medRxiv
Show abstract

This study explores the potential of state-of-the-art large language models (LLM) to scaffold type 1 diabetes management by automating the analysis of multimodal diabetes device data, including blood glucose, carbohydrate, and insulin. By conducting a series of empirically grounded data analysis tasks, such as detecting glycemic episodes, clustering similar episodes into patterns, identifying counterfactual days, and performing visual data analysis, we assess whether models like ChatGPT 4o, Claude 3.5 Sonnet, and Gemini Advanced can offer meaningful insights from diabetes data. Our findings show that ChatGPT 4o demonstrates strong potential in accurately interpreting data in the context of specific glycemic episodes, identifying glycemic patterns, and analyzing patterns. However, limitations in handling edge cases and visual reasoning tasks highlight areas for future development. Using LLMs to automate data analysis tasks and generate narrative summaries could scaffold clinical decision-making in diabetes management, which could make frequent data review feasible for improved patient outcomes.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.