LFMD: a new likelihood-based method to detect low-frequency mutations without molecular tags
Ye, R.; Ruan, J.; Zhuang, X.; Qi, Y.; An, Y.; Xu, J.; Mak, T.; Liu, X.; Yang, H.; Xu, X.; Baum, L.; Nie, C.; Sham, P. C.
Show abstract
As next-generation sequencing (NGS) and liquid biopsy become more prevalent in research and in the clinic, there is an increasing need for better methods to reduce cost and improve sensitivity and specificity of low-frequency mutation detection (where the Alternative Allele Frequency, or AAF, is less than 1%). Here we propose a likelihood-based approach, called Low-Frequency Mutation Detector (LFMD), which combines the advantages of duplex sequencing (DS) and the bottleneck sequencing system (BotSeqS) to maximize the utilization of duplicate reads. Compared with the existing state-of-the-art methods, DS, Du Novo, UMI-tools, and Unified Consensus Maker, our method achieves higher sensitivity, higher specificity (< 4 x 10-10 errors per base sequenced) and lower cost (reduced by ~70% at best) without involving additional experimental steps, customized adapters or molecular tags. LFMD is useful in areas where high precision is required, such as drug resistance prediction and cancer screening. As an example of LFMDs applications, mitochondrial heterogeneity analysis of 28 human brain samples across different stages of Alzheimers Disease (AD) showed that the canonical oxidative damage related mutations, C:G>A:T, are significantly increased in the mid-stage group. This is consistent with the Mitochondrial Free Radical Theory of Aging, suggesting that AD may be linked to the aging of brain cells induced by oxidative damage.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- iCOMIC: a graphical interface-driven bioinformatics pipeline for analyzing cancer omics data 94%
- FLYNC: A Machine Learning-Driven Framework for Discovering Long Non-Coding RNAs in Drosophila melanogaster 94%
- Kmerator Suite: design of specific k-mer signatures andautomatic metadata discovery in large RNA-Seq datasets. 94%
Similar papers in this journal
- HISAT-3N: a rapid and accurate three-nucleotide sequence aligner 96%
- Haplocheck: Phylogeny-based Contamination Detection in Mitochondrial and Whole-Genome Sequencing Studies 94%
- Ultra-low input single tube linked-read library method enables short-read NGS systems to generate highly accurate and economical long-range sequencing information for de novo genome assembly and haplotype phasing 94%
Similar papers in this journal
- 3rd-ChimeraMiner: A pipeline for integrated analysis of whole genome amplification generated chimeric sequences using long-read sequencing 95%
- Accurate Identification of Extrachromosomal Circular DNA from Long-read Sequences. 95%
- Comparing full variation profile analysis with the conventional consensus method in SARS-CoV-2 phylogeny 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.