Back

Clinical Trial Data Sharing: A Cross-Sectional Study of Outcomes Associated with Two NIH Models

Rowhani-Farid, A.; Egilman, A. C.; Zhang, A. D.; Gross, C. P.; Krumholz, H. M.; Ross, J. S.

2021-09-13 health policy
10.1101/2021.09.10.21263404 medRxiv
Show abstract

BackgroundThe impact and value of clinical trial data sharing, including the number and quality of publications that result from shared data - "shared data publications" - may differ depending on the data sharing model used. MethodsWe characterized the outcomes associated with two data sharing models previously used by Institutes of the U.S. National Institutes of Health (NIH): NHLBIs centralized model, which uses a repository to manage data sharing requests, and NCIs decentralized model, which entrusted research groups to independently manage data sharing requests. We identified trials completed in 2010 that met NIH data sharing criteria and matched studies sponsored by each Institute based on cost or size, determining whether trial data were shared and the frequency of shared data publications. ResultsWe identified 14 NHLBI-funded trials and 48 NCI-funded trials that met NIH data sharing criteria. We matched 14 NCI-funded trials to the 14 NHLBI-funded trials; among these, 4 NHLBI-sponsored trials (29%) and 2 NCI-sponsored trials (14%) shared data. From the 2 NCI-sponsored trials sharing data, we identified 2 shared data publications, one per trial, both of which were meta-analyses. From the 4 NHLBI-sponsored trials sharing data, we identified 7 shared data publications, all using data from 1 trial, 5 of which were pooled analyses and 2 reported secondary outcomes. ConclusionWhen characterizing the outcomes associated with two NIH data sharing models, both the NHLBI and the NCI models resulted in only 21% of trials sharing data and few shared data publications. There are opportunities to optimize clinical trial data sharing efforts both to enhance clinical trial data sharing and increase the number of shared data publications.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.