Resource Profile: The Regenstrief Institute COVID-19 Research Data Commons (CoRDaCo)
Allen, K. S.; Zidan, N.; Dey, V.; Mendonca, E.; Grannis, S.; Kasturi, S.; Khan, B.; Zappone, S. R.; Haggstrom, D.; Ruppert, L.; Schleyer, T.; Ning, X.; Embi, P.; Tachinardi, U.
Show abstract
The primary objective of the COVID-19 Research Data Commons (CoRDaCo) is to provide broad and efficient access to a large corpus of clinical data related to COVID-19 in Indiana, facilitating research and discovery. This curated collection of data elements provides information on a significant portion of COVID-19 positive patients in the State from the beginning of the pandemic, as well as two years of health information prior its onset. CoRDaCo combines data from multiple sources, including clinical data from a large, regional health information exchange, clinical data repositories of two health systems, and state laboratory reporting and vital records, as well as geographic-based social variables. Clinical data cover information such as healthcare encounters, vital measurements, laboratory orders and results, medications, diagnoses, the Charlson Comorbidity Index and Pediatric Early Warning Score, COVID-19 vaccinations, mechanical ventilation, restraint use, intensive care unit and ICU and hospital lengths of stay, and mortality. Interested researchers can visit ridata.org or email askrds@regenstrief.org to discuss access to CoRDaCo. Key FeaturesO_LICoRDaCo includes patient-level data on diagnosis and treatment, healthcare utilization, outcomes, and demographics. The level of detail available for each patient varies depending on the source of the clinical data. C_LIO_LICoRDaCo uses geographic identifiers to link patient-specific data to area-level social factors, such as census variables and social deprivation indices. C_LIO_LIAs of 4/30/21, the CoRDaCo cohort consists of over 776,000 cases, including granular data on over 15,000 patients who were admitted to an intensive care unit, and over 1,362,000 COVID-19-negative controls. Data is currently refreshed two times per month. C_LIO_LIThe most prevalent comorbidities in the data set include hypertension, diabetes, chronic pulmonary disease, renal disease, cancer, and congestive heart failure. C_LI
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- An Atomic Approach to the Design and Implementation of a Research Data Warehouse 95%
- Development and Validation of Phenotype Classifiers across Multiple Sites in the Observational Health Sciences and Informatics (OHDSI) Network 94%
- Increasing Trust in Real-World Evidence Through Evaluation of Observational Data Quality 94%
Similar papers in this journal
Similar papers in this journal
- LinkR: an open source, low-code and collaborative data science platform for healthcare data analysis and visualization 93%
- Machine Learning Directed Interventions Associate with Decreased Hospitalization Rates in Hemodialysis Patients 92%
- Predicting Prognosis in COVID-19 Patients using Machine Learning and Readily Available Clinical Data 91%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.