CHANGER: A novel genetic reference from 902 unrelated Chilean exomes to enhance South American medical genetics and research.
Gonzalez, E.; Villaman, C.; Rebolledo-Jaramillo, B.; Hernandez, C. F.; Bustos, B. I.; Berrios, D.; Moreno, G.; Posey, J. E.; Lupski, J. R.; Poli, C.; Calderon, J. F.; Munoz-Venturelli, P.; Fernandez, M. I.; Lecaros, J. A.; Armisen, R.; Repetto, G. M.; Perez-Palma, E.
Show abstract
BackgroundGenetic studies have disproportionately focused on populations of European ancestry, limiting the generalizability of allele-frequency references and genetic associations to underrepresented groups, including South American populations. This gap is particularly relevant for rare diseases and cancer, where accurate variant interpretation depends in part on appropriate population context. In addition, population-specific haplotype structure influence genome-wide association analyses and the portability of polygenic scores across ancestries. By focusing on the Chilean population, our work aims to bridge these gaps, providing more accurate reference data for genetic research and enhancing diagnostic and therapeutic strategies for underrepresented communities. MethodsWe aggregated exomes from seven Chilean cohorts and processed samples using a standardized best practices workflow, including population-scale quality control and relatedness filtering. Aggregated variants were annotated with VEP (including LOFTEE, REVEL, AlphaMissense, EVE, and CADD) and clinically classified using InterVar. Population structure was assessed with PCA/admixture. We compared overlap and allele frequencies against external references, performed gene-level variant burden analysis, and constructed phased haplotypes for ancestry association using linear regression with Bonferroni correction. Finally, we developed a dashboard with R shiny framework to enable gene-wise exploration of annotated variants. ResultsWe present CHANGER (Chilean Aggregated National Genomics Resource), a comprehensive aggregation of 902 unrelated Chilean exomes designed to create a detailed reference of genetic variation in Chile. By incorporating data from multiple cohorts, we identified 774,110 unique genetic variants, with an average of 42,601 aggregated variants per individual. We identified 132,363 novel variants, of which 31,470 were common within our cohort (frequency > 1 %). We provide variant annotation and direct comparisons with other publicly available general population references. Beyond variant discovery, CHANGER supports gene-level burden scans (61 Missense-Damaging depleted genes; 2 Loss-of-Function enriched genes) and a haplotype resource for ancestry association (51,171 haplotypes), enabling downstream interpretation, improved imputation, and more powerful genome-wide association analyses in Chileans. ConclusionsCHANGER provides a valuable resource for genetic research and a reference for local variant interpretation. Importantly, CHANGER follows a high-quality methodological baseline and an open, FAIR-aligned infrastructure designed to grow, increasing its value for discovery, imputation, and equitable clinical interpretation over time.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A phenome-wide association study identifies effects of copy number variation of VNTRs and multicopy genes on multiple human traits 95%
- Characterization of exome variants and their metabolic impact in 6,716 American Indians from Southwest US 95%
- Advancing long-read nanopore genome assembly and accurate variant calling for rare disease detection 95%
Similar papers in this journal
- Defining and Reducing Variant Classification Disparities 95%
- Leveraging genomic diversity for discovery in an EHR-linked biobank: the UCLA ATLAS Community Health Initiative 94%
- Systematic analysis of genetic and phenotypic characteristics reveals antisense oligonucleotide therapy potential for one-third of neurodevelopmental disorders 94%
Similar papers in this journal
- Systematic single-variant and gene-based association testing of 3,700 phenotypes in 281,850 UK Biobank exomes 96%
- Targeted long-read sequencing facilitates phased diploid assembly and genotyping of the human T cell receptor alpha, delta and beta loci 94%
- The UCLA ATLAS Community Health Initiative: promoting precision health research in a diverse biobank 94%
Similar papers in this journal
- Evaluation of imputation performance of multiple reference panels in a Pakistani population 96%
- A reference panel for linkage disequilibrium and genotype imputation using whole-genome sequencing data from 2,680 participants across India 94%
- Long-read genome sequencing for the diagnosis of neurodevelopmental disorders 94%
Similar papers in this journal
- Discovery of A Polymorphic Gene Fusion via Bottom-Up Chimeric RNA Prediction 96%
- Genome-wide characterization of human minisatellite VNTRs: population-specific alleles and gene expression differences 96%
- Long-read whole-genome sequencing-based concurrent haplotyping and aneuploidy profiling of single cells 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.