Geographic assessment of cancer genome profiling studies
Carrio Cordo, P.; Acheson, E.; Huang, Q.; Baudis, M.
Show abstract
Cancers arise from the accumulation of somatic genome mutations, which can be influenced by inherited genomic variants and external factors such as environmental or lifestyle-related exposure. Due to the heterogeneity of cancers, precise information about the genomic composition of germline and malignant tissues has to be correlated with morphological, clinical and extrinsic features to advance medical knowledge and treatment options. With global differences in cancer frequencies and disease types, geographic data is of importance to understand the interplay between genetic ancestry and environmental influence in cancer incidence, progression and treatment outcome. In this study, we analysed the current landscape of oncogenomic screening publications for geographic information content and quality, to address underrepresented study populations and thereby to fill prominent gaps in our understanding of interactions between somatic variations, population genetics and environmental factors in oncogenesis. We conclude that while the use of proxy derived geographic annotations can be useful for coarse-grained associations, the study of geo-correlated factors in cancer causation and progression will benefit from standardized geographic provenance annotations. Additionally, publication derived geographic provenance data allowed us to highlight stark inequality in the geographies of cancer genome profiling, with a near lack of sizeable studies from Africa and other large regions.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Comprehensive cancer-oriented biobanking resource of human samples for studies of post-zygotic genetic variation involved in cancer predisposition 93%
- The Rise of Open Data Practices Among Bioscientists at the University of Edinburgh 92%
- Novel candidates of pathogenic variants of the BRCA1 and BRCA2 genes in a 3,552 Japanese whole-genome sequence dataset (3.5KJPNv2) 92%
Similar papers in this journal
- GEGA (Gallus Enriched Gene Annotation): an online tool providing genomics and functional information across 47 tissues for a chicken gene-enriched atlas gathering Ensembl & Refseq genome annotations 93%
- CRUX, a platform for visualising, exploring and analysing cancer genome cohort data 93%
- iCOMIC: a graphical interface-driven bioinformatics pipeline for analyzing cancer omics data 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.