Monitoring of Open Science practices: a survey of 10 major medical journals
VINATIER, C.; Dewi, A. P. M.; Freyermuth, G.; Millour, M.; Scheidecker, M.; Arnault, F.-J.; Acher, M.; Nachev, V.; Weissgerber, T. L.; DeVito, N. J.; Dumont, G.; Gopalakrishna, G.; Le Bartz Lyan, G.; Stegeman, I.; Leeflang, M. M. G.; Naudet, F.
Show abstract
ObjectiveTo evaluate open science policies of leading general medical journals and the extent to which open science practices are implemented and detectable using automated tools. DesignCross-sectional audit of journal policies and retrospective observational study of journal articles, with diagnostic accuracy validation of automated screening tools against manual data extraction. Setting: Annals of Internal Medicine, BMC Medicine, The BMJ, CMAJ, JAMA, JAMA Network Open, The Lancet, Nature Medicine, New England Journal of Medicine, and PLOS Medicine. Participants:Research articles published between 2020 and 2023 and randomised controlled trials (RCTs) published before and after policy changes regarding registration and data-sharing. Main outcome measures: Journal policies were assessed using the TOP2025 framework, which rates journals transparency and openness standards using the TOP Factor (maximum score 27). At the article level, thirteen core open science practices were examined, including trial registration, protocol sharing, and intention to share data. Nine validated automated tools were applied to detect these practices and compared with manual extraction of 312 articles performed in duplicate. Changes in transparency practices following policy updates were analyzed. ResultsTransparency policies varied considerably. TOP Factor scores ranged from 1 (NEJM) to 13 (PLOS Medicine), with many journals policies applying primarily to clinical trials rather than all research articles. Only one journal (BMC Medicine) proposed registered reports. At the article level, adoption of open science practices was highest in RCTs: trial registration (99% [95%-100%] vs 68% [55%-79%] in meta-analyses vs 15% [8%-27%] in other research), protocol sharing (95% [90%-98%] vs 67% [54%-78%] vs 19% [11%-32%]), and intention to share data (78% [67%-87%] vs 64% [49%-76%] vs 69% [55%-80%]). Performance across tools was very variable (F1 scores 0.16-1.00). Among all 15,624 research articles, tools detected registration in 20%, protocol sharing in 42%, and intention to share data in 48%, but these were generally underestimated. Policy changes such as mandatory registration or data-sharing requirements were associated with measurable improvements in the corresponding open science practices. ConclusionsLeading medical journals show only partial alignment with TOP2025, with stronger transparency observed for RCTs than for other study types. Triangulation with manual and automated assessments in articles confirms that key open science practices remain suboptimal especially for non-RCTs, highlighting the need for more robust journal policies. Policy changes are associated with observable improvements. Registration: OSF: https://doi.org/10.17605/OSF.IO/F2VW9
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- COVID-19-related research data availability and quality according to the FAIR principles: A meta-research study 97%
- The Rise of Open Data Practices Among Bioscientists at the University of Edinburgh 96%
- Introducing the EMPIRE Index: A novel, value-based metric framework to measure the impact of medical publications 95%
Similar papers in this journal
Similar papers in this journal
- The use of the Registered Reports format for publication of randomized clinical trials: a cross-sectional study 94%
- Results reporting for clinical trials led by medical universities and university hospitals in the Nordic countries was often missing or delayed 93%
- Large language models for conducting systematic reviews: on the rise, but not yet ready for use – a scoping review 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.