Back

Yeast rDNA as a benchmark for rDNAmine repeat analysis pipeline

Czarnocka-Cieciura, A. M.; Guminska, N.

2026-02-25 genomics
10.64898/2026.02.24.707673 bioRxiv
Show abstract

In this study, we introduce a novel approach for analysing long, repetitive genomic sequences. Our methods significantly advance research on rDNA polymorphism. First, we describe a technique for isolating high-molecular-weight DNA from individual chromosomes, enabling selective enrichment of sequencing libraries for extensive genomic regions of interest. Second, we present rDNAmine, a bioinformatic toolkit for capturing and examining large repetitive arrays in Oxford Nanopore sequencing data. This approach facilitates the study of polymorphisms within long repeats, bypassing traditional alignment-based methods and providing a more efficient and scalable solution for investigating repetitive regions. We demonstrate the effectiveness of our approach through the analysis of rDNA arrays in two yeast species, Saccharomyces cerevisiae and Candida albicans. In S. cerevisiae, rDNA arrays show limited polymorphism, while in C. albicans, we observe substantial variation in rDNA module size, with two distinct repeat populations within the array. These findings reveal species-specific differences in the structural organisation of rDNA loci, highlighting the diverse nature of tandem repeat architecture. The rDNAmine toolkit is broadly applicable to various organisms and repetitive genomic contexts, offering a versatile platform for studying repetitive sequences. Take AwayO_LIYeast rDNA serves as a benchmark to validate tools for analysing long repetitive sequences. C_LIO_LIA chromosome-specific DNA extraction method has been introduced to enable targeted enrichment of repetitive loci. C_LIO_LIThe rDNAmine pipeline was designed to analyse long tandem repeats from noisy long-read data without requiring global alignment. C_LI

Published in Yeast (predicted rank #5) · training set

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
G3
33 papers in training set
Top 0.1%
14.9%
2
Molecular Ecology Resources
171 papers in training set
Top 0.2%
9.5%
3
G3: Genes, Genomes, Genetics
252 papers in training set
Top 0.4%
9.5%
4
G3 Genes|Genomes|Genetics
351 papers in training set
Top 0.6%
7.8%
Yeast · published here
17 papers in training set
Top 0.1%
6.2%
6
Genome Research
468 papers in training set
Top 1%
4.3%
50% of probability mass above
7
Genome Biology and Evolution
338 papers in training set
Top 1%
4.0%
8
BMC Genomics
406 papers in training set
Top 2%
4.0%
9
Peer Community Journal
281 papers in training set
Top 2%
3.2%
10
Nucleic Acids Research
1281 papers in training set
Top 6%
3.2%
11
Genomics
64 papers in training set
Top 0.5%
2.4%
12
GENETICS
483 papers in training set
Top 2%
2.4%
13
Frontiers in Microbiology
427 papers in training set
Top 5%
2.1%
14
Microbial Genomics
225 papers in training set
Top 1%
1.9%
15
NAR Genomics and Bioinformatics
242 papers in training set
Top 3%
1.5%
16
Molecular Biology and Evolution
542 papers in training set
Top 4%
1.3%
17
eLife
5828 papers in training set
Top 58%
1.1%
18
Bioinformatics
1204 papers in training set
Top 8%
1.1%
19
Frontiers in Fungal Biology
10 papers in training set
Top 0.1%
1.0%
20
Computational and Structural Biotechnology Journal
242 papers in training set
Top 6%
1.0%
21
BMC Biology
265 papers in training set
Top 4%
1.0%
22
PeerJ
308 papers in training set
Top 12%
0.8%
23
BMC Bioinformatics
457 papers in training set
Top 6%
0.8%
24
Scientific Reports
3612 papers in training set
Top 75%
0.8%
25
mSystems
394 papers in training set
Top 6%
0.8%