Tracing the evolutionary path of the CCR5delta32 deletion via ancient and modern genomes
Ravn, K.; Cobuccio, L.; Muktupavela, R. A.; Meisner, J. M.; Benros, M. E.; Korneliussen, T. S.; Sikora, M.; Willerslev, E.; Allentoft, M. E.; Irving-Pease, E. K.; Racimo, F.; Rasmussen, S.
Show abstract
The chemokine receptor variant CCR5delta32 is linked to HIV-1 infection resistance and other pathological conditions. In European populations, the allele frequency ranges from 10-16%, and its evolution has been extensively debated throughout the years. We provide a detailed perspective of the evolutionary history of the deletion through time and space. We discovered that the CCR5delta32 allele arose on a pre-existing haplotype consisting of 84 variants. Using this information, we developed a haplotype-aware probabilistic model to screen for this deletion across 860 low-coverage ancient genomes and we found evidence that CCR5delta32 arose at least 7,000 years BP, with a likely origin somewhere in the Western Eurasian Steppe region. We further show evidence that the CCR5delta32 haplotype underwent positive selection between 7,000-2,000 BP in Western Eurasia and that the presence of the haplotype in Latin America can be explained by post-Columbian genetic exchanges. Finally, we point to new complex CCR5delta32 genotype-haplotype-phenotype relationships, which demand consideration when targeting the CCR5 receptor for therapeutic strategies.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genetic adaptation to pathogens and increased risk of inflammatory disorders in post-Neolithic Europe 97%
- Ancient genomes illuminate Eastern Arabian population history and adaptation against malaria 96%
- Genomic Insights into the Demographic History and Local Adaptation of Wild Boars Across Eurasia 96%
Similar papers in this journal
- Accurate rare variant phasing of whole-genome and whole-exome sequencing data in the UK Biobank 95%
- Leveraging fine-mapping and non-European training data to improve trans-ethnic polygenic risk scores 95%
- Genotyping sequence-resolved copy number variationusing pangenomes reveals paralog-specific global diversityand expression divergence of duplicated genes 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.