Long-read sequencing reveals novel mitochondrial genome variants undetected by short-read sequencing in Korean population
Kim, H. J.; Kim, S. M.; Yoon, K.; Kim, B.-J.; Jin, H. J.; Kim, Y. J.
Show abstract
While long-read sequencing technologies (e.g., PacBio Revio, ONT) have revolutionized high-quality genome assembly for the human pangenome, mitochondrial genome (mtDNA) analysis still largely relies on short-read and Sanger sequencing. However, short-read sequencing often lacks the resolution required to resolve complex variations due to the unique features of mtDNA, such as high mutation rates and repetitive homopolymeric regions, which frequently lead to alignment artifacts and mapping ambiguities. To address this, we evaluated whether applying long-read sequencing to mtDNA improves analytical quality in empirical data. Through comprehensive bioinformatics analyses, we compared the performance of long-read sequencing against short-read sequencing and microarrays. Our results revealed that long-read sequencing detected the highest number of variants (n = 533), significantly outperforming both short-read sequencing (n = 525) and microarrays (n = 49). Notably, both sequencing methods provided significantly higher resolution in haplogroup assignment compared to microarrays in terms of phylogenetic depth (p < 0.05). Long-read sequencing demonstrated superior detection power, particularly for InDels. We identified two novel non-synonymous variants, including a unique InDel detected exclusively by long-read sequencing. Protein modeling and stability analysis validated that this InDel causes structural instability (RMSD > 2.0 [A],{Delta}{Delta} G = -45.21 kcal/mol). Furthermore, we confirmed that this novel InDel is shared among haplogroup A samples in both the 1000 Genomes Project ONT dataset and the Korean population, highlighting the practical implications of long-read sequencing for molecular biology and population genetics.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Long read mitochondrial genome sequencing using Cas9-guided adaptor ligation 96%
- Reanalysis of mtDNA mutations of human primordial germ cells (PGCs) reveals NUMT contamination and suggests that selection in PGCs may be positive 91%
- Presence and Transmission of Mitochondrial Heteroplasmic Mutations in Human Populations of European and African Ancestry 90%
Similar papers in this journal
- Structural variation of the malaria-associated human glycophorin A-B-E region 93%
- Systematic benchmark of state-of-the-art variant calling pipelines identifies major factors affecting accuracy of coding sequence variant discovery 93%
- A new method for long-read sequencing of animal mitochondrial genomes: application to the identification of equine mitochondrial DNA variants 93%
Similar papers in this journal
- Low-pass sequencing plus imputation using avidity sequencing displays comparable imputation accuracy to sequencing by synthesis while reducing duplicates 93%
- A High-quality Oxford Nanopore Assembly of the Hourglass Dolphin (Lagenorhynchus cruciger) Genome 93%
- First Whole-Genome Assembly of the Galapagos Petrel (Pterodroma phaeopygia) Using Oxford Nanopore Sequencing to Advance Conservation Genomics in a Critically Endangered Seabird 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.