Detecting oncogenic selection through biased allele retention in The Cancer Genome Atlas
Luft, J.; Young, R. S.; Meynert, A. M.; Taylor, M. S.
Show abstract
Background The loss of genetic diversity in segments over a genome (loss-of-heterozygosity, LOH) is a common occurrence in many types of cancer. By analysing patterns of preferential allelic retention during LOH in approximately 10,000 cancer samples from The Cancer Genome Atlas (TCGA), we sought to systematically identify genetic polymorphisms currently segregating in the human population that are preferentially selected for, or against during cancer development.Results Experimental batch effects and cross-sample contamination were found to be substantial confounders in this widely used and well studied dataset. To mitigate these we developed a generally applicable classifier (GenomeArtiFinder) to quantify contamination and other abnormalities. We provide these results as a resource to aid further analysis of TCGA whole exome sequencing data. In total, 1,678 pairs of samples (14.7%) were found to be contaminated or affected by systematic experimental error. After filtering, our analysis of LOH revealed an overall trend for biased retention of cancer-associated risk alleles previously identified by genome wide association studies. Analysis of predicted damaging germline variants identified highly significant oncogenic selection for recessive tumour suppressor alleles. These are enriched for biological pathways involved in genome maintenance and stability.Conclusions Our results identified predicted damaging germline variants in genes responsible for the repair of DNA strand breaks and homologous repair as the most common targets of allele biased LOH. This suggests a ratchet-like process where heterozygous germline mutations in these genes reduce the efficacy of DNA double-strand break repair, increasing the likelihood of a second hit at the locus removing the wild-type allele and triggering an oncogenic mutator phenotype.Competing Interest StatementThe authors have declared no competing interest.View Full Text
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Cross-dataset pan-cancer detection: Correlating cell-free DNA fragment coverage with open chromatin sites across cell types 97%
- Extreme intratumour heterogeneity and driver evolution in mismatch repair deficient gastro-oesophageal cancer 97%
- The DiffInvex evolutionary model for conditional somatic selection identifies chemotherapy resistance genes in 10,000 cancer genomes 97%
Similar papers in this journal
- Long-read sequencing of diagnosis and post-therapy medulloblastoma reveals complex rearrangement patterns and epigenetic signatures 96%
- Joint estimation and imputation of variant functional effects using high throughput assay data 95%
- Polymorphic short tandem repeats make widespread contributions to blood and serum traits 95%
Similar papers in this journal
- RNA allelic frequencies of somatic mutations encode substantial functional information in cancers 97%
- Disruption of metazoan gene regulatory networks in cancer alters the balance of co-expression between genes of unicellular and multicellular origins 96%
- SVFX: a machine-learning framework to quantify the pathogenicity of structural variants 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.