Back

EvidenceTriangulator: A Large Language Model Approach to Synthesizing Causal Evidence across Study Designs

Shi, X.; Zhao, W.; Yang, C.; Du, J.

2024-03-19 health informatics
10.1101/2024.03.18.24304457 medRxiv
Show abstract

Health strategies increasingly emphasize both behavioral and biomedical interventions, yet the complex and often contradictory guidance on diet, behavior, and health outcomes complicates evidence-based decision-making. Evidence triangulation across diverse study designs is essential for establishing causality, but scalable, automated methods for achieving this are lacking. In this study, we assess the performance of large language models (LLMs) in extracting both ontological and methodological information from scientific literature to automate evidence triangulation. A two-step extraction approach--focusing on cause-effect concepts first, followed by relation extraction--outperformed a one-step method, particularly in identifying effect direction and statistical significance. Using salt intake and blood pressure as a case study, we calculated the Convergeny of Evidence (CoE) and Level of Evidence (LoE), finding a trending excitatory effect of salt on hypertension risk, with a moderate LoE. This approach complements traditional meta-analyses by integrating evidence across study designs, thereby facilitating more comprehensive assessments of public health recommendations.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.