Real-World Evidence BRIDGE: a tool to connect protocol with code programming
Cid Royo, A.; Elbers, R.; Weibel, D.; Hoxhaj, V.; Kurkcuoglu, Z.; Sturkenboom, M. C.; Vaz, T. A.; Andaur Navarro, C. L.
Show abstract
ObjectiveO_ST_ABSMethodsC_ST_ABSSeveral statistical analysis plans (SAP) from the Vaccine Monitoring Collaboration for Europe (VAC4EU) were analyzed to identify the study design sections and specifications for programming RWE studies based on multi-databases standardized to common data models. We envisioned a metadata schema that transforms the epidemiologists knowledge into a machine-readable format. This machine-readable metadata schema must also contain the different study sections, code lists, and time anchoring specified in the SAPs. Further desired attributes are adaptability and user-friendliness. ResultsWe developed RWE-BRIDGE, a metadata schema with a star-schema model divided into four study design sections with 12 tables: Study Variable Definition with two tables, Cohort Definition with two tables, Post-Exposure Outcome Analysis with one table, and Data Retrieval with seven tables. We provide examples and a step-by-step guide to populate this metadata schema. In addition, we provide a Shiny app that checks the several tables proposed in this metadata strategy. RWE-BRIDGE is available at https://github.com/UMC-Utrecht-RWE/RWE-BRIDGE. DiscussionThe RWE-BRIDGE has been designed to support the translation of study design sections from statistical analysis plans into analytical pipelines, facilitating collaboration and transparency between lead researchers and scientific programmers and reducing hard coding and repetition. This metadata schema strategy is flexible by supporting different common data models and programming languages, and it is adaptable to the specific needs of each SAP by adding further tables or fields, if necessary. Modified versions of the RWE-BRIGE have been applied in several RWE studies within the VAC4EU ecosystem. ConclusionThe RWE-BRIDGE offers a systematic approach to detailing what type of variables, time anchoring, and algorithms are required for a specific RWE study. Applying this metadata schema can facilitate the communication between epidemiologists and programmers in a transparent manner.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- An Ontology-based Approach to Guide and Document Variable and Data Source Selection and Data Integration Process to Support Integrative Data Analysis in Cancer Outcomes Research 93%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 93%
- Evaluating Semantic Similarity Methods for Comparison of Text-derived Phenotype Profiles 93%
Similar papers in this journal
- Advancing data science in drug development through an innovative computational framework for data sharing and statistical analysis 93%
- Quantitative bias analysis for mismeasured variables in health research: a review of software tools 93%
- Scalable information extraction from free text electronic health records using large language models 92%
Similar papers in this journal
- Transforming Estonian health data to the Observational Medical Outcomes Partnership (OMOP) Common Data Model: lessons learned 96%
- Trajectories: a framework for detecting temporal clinical event sequences from health data standardized to the OMOP Common Data Model 95%
- A Simple Electronic Medical Record System Designed for Research 95%
Similar papers in this journal
- De-novo FAIRification via an Electronic Data Capture system by automated transformation of filled electronic Case Report Forms into machine-readable data 96%
- EHR-QC: A streamlined pipeline for automated electronic health records standardisation and preprocessing to predict clinical outcomes 95%
- Using computable knowledge mined from the literature to elucidate confounders for EHR-based pharmacovigilance 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.