The HuBMAP Framework for Advancing Data FAIRness
Fisher, S. A.; Hardi, J.; Morgan, R.; Nordgren, E.; Kant, P. M.; Honick, B.; Rosario, J.; O'Connor, M. J.; Blood, P. D.; HuBMAP DCWG Members, ; Silverstein, J. C.; Musen, M. A.
Show abstract
Since publication of the FAIR Guiding Principles in 2016, the scientific community has increasingly sought to make experimental data findable, accessible, interoperable, and reusable. Operationalizing the FAIR principles in routine scientific workflows remains challenging without a standardized, workable infrastructure. With over 10,000 datasets from over 40 institutions, spanning more than 50 diverse assay types ranging from single-cell sequencing technologies to 2D and 3D spatial omics, the U.S. National Institutes of Health (NIH) Human Bio-Molecular Atlas Program (HuBMAP) consortium has been ideally situated to create a FAIR ecosystem. With the goal of achieving data "FAIRness," HuBMAP developed and implemented well-defined, community-endorsed metadata reporting standards across the research lifecycle. These reporting standards include detailed schemas, harmonized across a multitude of assays, that define the metadata associated with a dataset and the organization of the corresponding data files. These standards ensure documentation of the data collection process, of the data themselves, and of the manner in which the data are packaged for sharing, while remaining compliant with the Health Insurance Portability and Accountability Act (HIPAA). The use of these reporting standards, in tandem with technology to foster adherence, allows HuBMAP to fulfill its goal of generating FAIR data for open dissemination through its Data Portal and Human Reference Atlas. The procedures and simple workflow adopted by HuBMAP investigators serve as a model for other scientific communities aiming to maximize the value of varied datasets addressing a shared research question. The HuBMAP end-to-end, metadata-centered workflow has been replicated and enhanced by the NIH Cellular Senescence Network (SenNet) consortium and is readily available through open-source technology for others to utilize.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- DeepSpaceDB: a spatial transcriptomics atlas for interactive in-depth analysis of tissues and tissue microenvironments 94%
- The Neuroscience Multi-Omic Archive: A BRAIN Initiative resource for single-cell transcriptomic and epigenomic data from the mammalian brain 93%
- Datanator: an integrated database of molecular data for quantitatively modeling cellular behavior 93%
Similar papers in this journal
Similar papers in this journal
- Sharing Data from the Human Tumor Atlas Network through Standards, Infrastructure, and Community Engagement 96%
- Human BioMolecular Atlas Program (HuBMAP): 3D Human Reference Atlas Construction and Usage 96%
- gEAR: gene Expression Analysis Resource portal for community-driven, multi-omic data exploration 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.