Exome-wide association studies discover germline mutation patterns and identify high-risk populations in human cancers
Shen, S.; Jiang, Y.; Wang, G.; Li, H.; You, D.; Duan, W.; Zhang, R.; Wei, Y.; Shen, H.; Hu, Z.; Christiani, D.; Zhao, Y.; Chen, F.
Show abstract
Genome-wide association studies have discovered numerous common variants associated with human cancers. However, the contribution of exome-wide rare variants to cancers remains largely unexplored, especially for the protein-coding variants. The UK Biobank provides detailed cancer follow-up information linked to whole-exome sequencing (WES) for approximately 450,000 participants, offering an unprecedented opportunity to evaluate the effect of exome variation on pan-cancer. Here, we performed exome-wide association studies (ExWAS) based on single variant levels and gene levels to detect their associations across 20 primary cancer types in the discovery set (WES-300k, N = 284,456) and replication set (WES-150k, N = 143,478), separately. The ExWAS detected 143 independent variants at variant-level and 49 genes at gene-level, while nine variants and eight genes were shared across cancers. In the cross-trait meta-analysis, we identified 239 additional independent pleiotropic variants, mapping to the genes which were functional through trans-omics analyses in transcriptomics and proteomics. Further, we developed exome-wide risk scores (ERS) to identify high-risk populations based on rare variants with minor allele frequency (MAF) < 0.05. The ERS had satisfactory performance in cancer risk stratification, especially for the extremely high-risk persons (top 5% ERS) that were frequently risk allele carriers. The ERS (median C-index (IQR): 0.655 (0.636-0.667)) outperforms the traditional polygenic risk score (PRS) (median C-index (IQR): 0.585 (0.572-0.614)) for discrimination in the replication set. Our findings offer further insight into the genetic architecture of human exomes for cancer susceptibility.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Cross-dataset pan-cancer detection: Correlating cell-free DNA fragment coverage with open chromatin sites across cell types 95%
- Risk factors for eight common cancers revealed from a phenome-wide Mendelian randomisation analysis of 378,142 cases and 485,715 controls 95%
- Pan-cancer analysis demonstrates that integrating polygenic risk scores with modifiable risk factors improves risk prediction 95%
Similar papers in this journal
- Pan-cancer proteogenomic landscape of whole-genome doubling reveals putative therapeutic targets in various cancer types 95%
- Multi-omics consensus ensemble refines the classification of muscle-invasive bladder cancer with stratified prognosis, tumour microenvironment and distinct sensitivity to frontline therapies 93%
- Genomic profiling of cell lines reveals hidden research bias and caveats 92%
Similar papers in this journal
- Recurrent disruption of tumour suppressor genes in cancer by somatic mutations in cleavage and polyadenylation signals 94%
- Unveiling the influence of tumor and immune signatures on immune checkpoint therapy in advanced lung cancer 94%
- Multi-gradient Permutation Survival Analysis Identifies Mitosis and Immune Signatures Steadily Associated with Cancer Patient Prognosis 94%
Similar papers in this journal
- Hereditary haemochromatosis beyond liver cancer: increased risk of prostate cancer during an 11-year follow-up 91%
- Genetic analysis of functional rare germline variants across 9 cancer types from the DiscovEHR study 91%
- Early Lung Cancer Detection Using Nucleotide Transition Probabilities in plasma cell-free DNA 91%
Similar papers in this journal
- A joint transcriptome-wide association study across multiple tissues identifies new candidate susceptibility genes for breast cancer 95%
- Cell type-specific meQTL extends melanoma GWAS annotation beyond eQTL and informs melanocyte gene regulatory mechanisms 94%
- Brain eQTLs of European, African American, and Asian ancestry improve interpretation of schizophrenia GWAS 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.