Back

A multi-ancestry genetic reference for the Quebec population

McClelland, P.; Femerling, G.; Laflamme, R.; Mejia-Garcia, A.; Dehkordi, M. S.; Xiao, H.; Diaz-Papkovich, A.; Pelletier, J.; Grenier, J.-C.; Lo, K. S.; Anderson-Trocme, L.; Bellavance, J.; Chapdelaine, V.; Gagnon, G.; De Mori, A.; Martinez, G.; Mohler, K.; de Malliard, T.; Labbe, C.; Labrecque, M.; Montpetit, A.; Theroux, J.-F.; Zhou, H.; Girard, S. L.; Hussin, J. G.; Laberge, A.-M.; Bherer, C.; Tetreault, M.; Gagliano Taliun, S. A.; Taliun, D.; Gravel, S.; Lettre, G.

2025-05-16 genetic and genomic medicine
10.1101/2025.05.14.25327536 medRxiv
Show abstract

While international efforts have characterized genetic variation in millions of individuals, the interplay of environmental, social, cultural, and genetic factors is poorly understood for most worldwide populations. The province of Quebec in Canada has been the site of numerous genetic studies, often focusing on individual Mendelian diseases in founder sub-populations. Here, we profiled and analyzed genome-wide genotyped variation in 29,337 Quebec residents from the large population-based cohort CARTaGENE (CaG), including rich phenotype and environmental data. We also sequenced the whole-genome of 2,173 CaG participants, including 163 and 132 individuals with grandparents born in Haiti and Morocco, respectively. We use this genetic information to gain insight into Quebecs demography and to help interpret the potential significance of variants identified in clinically important genes. We built an imputation panel by phasing the CaG whole-genome sequence data and showed, using genome-wide association studies (GWAS), how it improves the discovery of phenotype-genotype associations in this population. We provide allele frequency information and GWAS results through dedicated and publicly available websites. The genetic data, paired with phenotypic and environmental information, is also available for research use upon scientific and ethical review.

Published in Nature Communications (predicted rank #2) · training set

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.