The Consortium for Genomic Diversity, Ancestry, and Health in Colombia (CODIGO): building local capacity in genomics, bioinformatics, and precision medicine
Marino-Ramirez, L.; Sharma, S.; Hamilton, J. M.; Nguyen, T. L.; Gupta, S.; Natarajan, A. V.; Nagar, S. D.; Menuey, J. L.; Chen, W.-A.; Sanchez-Gomez, A.; Satizabal-Soto, J. M.; Martinez, B.; Marrugo, J.; Medina-Rivas, M.; Gallo, J. E.; Jordan, I. K.; Valderrama-Aguirre, A.
Show abstract
The Consortium for Genomic Diversity, Ancestry, and Health in Colombia (CODIGO) aims to build a community of Colombian researchers in support of local capacity in genomics, bioinformatics, and precision health. Here, we present the first CODIGO data release and the consortium web platform, including annotations for more than 95 million genetic variants from 1,441 samples representing 14 populations from across the country. CODIGO samples show a wide range of African (16.7%), European (50.6%), and Indigenous American (32.8%) genetic ancestry components, with five distinct ancestry clusters. Thousands of ancestry-enriched variants, with divergent allele frequencies across clusters, show pharmacogenomic and clinical genetic associations. Examples include African ancestry-enriched variants associated with fast metabolism of the immunosuppressive drug tacrolimus and malaria resistance and European ancestry-enriched variants associated with nicotine dependence and hereditary hemochromatosis. CODIGO reveals the nexus between ancestry and health in Colombia and underscores the utility of collaborative genome sequence analysis efforts.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Defining and Reducing Variant Classification Disparities 95%
- Multi-modal investigation of the schizophrenia-associated 3q29 genomic interval reveals global genetic diversity with unique haplotypes and segments that increase the risk for non-allelic homologous recombination 94%
- Genome-wide prediction of pathogenic gain- and loss-of-function variants from ensemble learning of diverse feature set 94%
Similar papers in this journal
- Widespread recessive effects on common diseases in a cohort of 44,000 British Pakistanis and Bangladeshis with high autozygosity 96%
- Summix: A method for detecting and adjusting for population structure in genetic summary data 96%
- Characterization of exome variants and their metabolic impact in 6,716 American Indians from Southwest US 95%
Similar papers in this journal
- Utilizing Non-Invasive Prenatal Test Sequencing Data Resource for Human Genetic Investigation 95%
- Genotyping and population structure of the China Kadoorie Biobank 95%
- The genetic and phenotypic correlates of neonatal Complement Component 3 and 4 protein concentrations with a focus on psychiatric and autoimmune disorders 94%
Similar papers in this journal
- A reference panel for linkage disequilibrium and genotype imputation using whole-genome sequencing data from 2,680 participants across India 95%
- Long-read genome sequencing for the diagnosis of neurodevelopmental disorders 95%
- Multivariate adaptive shrinkage improves cross-population transcriptome prediction for transcriptome-wide association studies in underrepresented populations 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.