Developing a GraphRAG-enabled local-LLM for Gestational Diabetes Mellitus.
Sharma, R.
Show abstract
This paper re-imagines a world of abundance in the treatment of chronic diseases such as Tpe 2 Diabetes. It asks: what if preventive and diagnostic remedies were widely made available across the world, informed by the latest medical research? As Proof-of-Concept of a proposed solution, the paper describes the development and validation of a local Large Language Models (local-LLMs) based on Graph-based Retrieval-Augmented Generation (GraphRAG) for managing Gestational Diabetes Mellitus (GDM). The research thus seeks new insights into optimizing GDM treatment through a knowledge graph architecture, contributing to a deeper understanding of how artificial intelligence can extend medical expertise to underserved populations globally. The study employs an agile, prototyping approach utilizing GraphRAG to enhance knowledge graphs by integrating retrieval-based and generative artificial intelligence techniques. Training data was from academic papers published between January 2000 and May 2024 using the Semantic Scholar API and analyzed by mapping complex associations within GDM management to create a comprehensive knowledge graph architecture. It is categorically stated that, since the primary research objective was to establish the feasibility of a GraphRAG local-LLM PoC, no human subjects nor actual patient datasets were used. Empirical results indicate that the GraphRAG-based Proof of Concept outperforms open-source LLMs such as ChatGPT, Claude, and BioMistral across key evaluation metrics. Specifically, GraphRAG achieves superior accuracy with BLEU scores of 0.99, Jaccard similarity of 0.98, and BERT scores of 0.98, offering significant implications for personalized medical insights that enhance diagnostic accuracy and treatment efficacy. This research offers a novel perspective on applying GraphRAG-enabled LLM technologies to GDM management, providing valuable insights that extend current understanding of AI applications in healthcare. The studys findings contribute to advancing the feasibility of GenAI for proactive GDM treatment and extending medical expertise to underserved populations globally.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Evaluating the impact on clinical task efficiency of a natural language processing algorithm for searching medical documents: Prospective crossover study 95%
- Transformative potential of Large Language Models in data mining on Electronic Health Records. 95%
- Assessment of Accuracy and Safety of LabTest Checker (LTC-AI) 93%
Similar papers in this journal
- Mining for Equitable Health: Assessing the Impact of Missing Data in Electronic Health Records 94%
- Graph-Based Clinical Recommender: Predicting Specialists Procedure Orders using Graph Representation Learning 94%
- EHR-QC: A streamlined pipeline for automated electronic health records standardisation and preprocessing to predict clinical outcomes 94%
Similar papers in this journal
Similar papers in this journal
- Collaborative intelligence in AI: Evaluating the performance of a council of AIs on the USMLE 94%
- Cardiology Knowledge Assessment of Retrieval-Augmented Open versus Proprietary Large Language Models 94%
- Development and preliminary testing of Health Equity Across the AI Lifecycle (HEAAL): A framework for healthcare delivery organizations to mitigate the risk of AI solutions worsening health inequities 94%
Similar papers in this journal
- Trajectories: a framework for detecting temporal clinical event sequences from health data standardized to the OMOP Common Data Model 95%
- Transforming Estonian health data to the Observational Medical Outcomes Partnership (OMOP) Common Data Model: lessons learned 94%
- A Simple Electronic Medical Record System Designed for Research 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.