The ground truth of the Data-Iceberg: Correct Meta-data
Caliskan, A.; Dangwal, S.; Dandekar, T.
Show abstract
Short summaryBiological molecular data such as sequence information increase so rapidly that detailed metadata, describing the process and conditions of data collection as well as proper labelling and typing of the data become ever more important to avoid mistakes and erroneous labeling. Starting from a striking example of wrong labelling of patient data recently published in Nature, we advocate measures to improve software metadata and controls in a timely manner to not rapidly loose quality in the ever-growing data flood.
Matching journals
The top 11 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- The current landscape and emerging challenges of benchmarking single-cell methods 91%
- Accurate Identification of Extrachromosomal Circular DNA from Long-read Sequences. 89%
- Accelerate the discovery of genetic variants in mitochondrial diseases with VIOLA: Variant PrIOritization using Latent space 89%
Similar papers in this journal
- CellPhoneDB v2.0: Inferring cell-cell communication from combined expression of multi-subunit receptor-ligand complexes 90%
- OCTAD: an open workplace for virtually screening therapeutics targeting precise cancer patient groups using gene expression features 90%
- Jointly Defining Cell Types from Multiple Single-Cell Datasets Using LIGER 88%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.