Back

Local causal discovery in epidemiology: an application to quantifying the effect of diabetes on severe liver fibrosis in patients with viral hepatitis

Loranchet, T.; Bystrova, D.; Burgat, P.; Bellet, J.; Bourliere, M.; Lusivika-Nzinga, C.; Nicol, J.; Boëlle, P.-Y.; Carrat, F.; Assaad, C. K.

2025-11-13 epidemiology
10.1101/2025.09.02.25334768 medRxiv
Show abstract

BackgroundEstimating the controlled direct effect (CDE) from observational data is challenging when the DAG is unknown. Causal discovery methods can infer a partially oriented DAG, enabling the identification of potential adjustment sets. We use a local causal discovery algorithm that focuses on the relevant portion of the graph, reducing assumptions and complexity compared to global methods. This approach is applied to a viral hepatitis cohort to estimate the CDE of diabetes on severe liver fibrosis. MethodsThe CDE of diabetes on liver fibrosis in patients with HBV or HCV was assessed using baseline data from the French ANRS CO22 HEPATHER cohort initiated in 2012. A local causal discovery algorithm, LocalPC-CDE, with bootstrap augmentation identified a robust adjustment set, retaining only variables minimally affected by sampling variability. The CDE was quantified as a causal odds ratio using logistic regression. ResultsCausal discovery included 20858 patients, with estimation performed on 8802 completecase observations. The algorithm identified an adjustment set of seven variables: geographical origin, age, hepatitis type, total cholesterol, HDL cholesterol, past alcohol consumption, blood glucose, and sex. The CDE of diabetes on severe fibrosis in viral hepatitis patients was significantly positive, with an estimated odds ratio of 2.03 (95% CI [1.78, 2.31]). ConclusionsAfter causal adjustment using a targeted, data-driven approach, diabetes retained a direct and statistically significant effect on liver fibrosis in patients with chronic viral hepatitis. This paper more generally introduces a methodological pipeline for local causal discovery when the underlying DAG is uncertain.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.