Back

Revisiting the base pair maximization approach for RNA secondary structure prediction with SQUARNA

Serdakov, M. D.; Bohdan, D. R.; Nikolaev, G. I.; Bujnicki, J. M.; Baulin, E. F.

2026-07-01 bioinformatics
10.64898/2026.06.30.735492 bioRxiv
Show abstract

Non-coding RNAs play diverse roles in a wide range of cellular processes, with their spatial structure being pivotal to their function. RNA secondary structure is a key determinant of its overall fold. Given the scarcity of experimentally determined RNA 3D structures, understanding secondary structure is vital for discerning RNA function. Currently, there is no universally effective solution for de novo RNA secondary structure prediction. Existing methods are becoming increasingly complex without marked improvements in accuracy and often overlook critical features such as pseudoknots and alternative folds. Here, we introduce SQUARNA, a new approach to de novo RNA secondary structure prediction that is suitable for both individual RNA analysis and large-scale structural searches. SQUARNA revisits the concept of base pair maximization and develops it into a stem maximization idea coupled with the widely used free energy minimization (MFE) framework. SQUARNA can predict alternative structures and handle pseudoknots of arbitrary complexity. Benchmarking shows that SQUARNA outperforms existing methods, including deep learning models, in both single-sequence and alignment-based RNA secondary structure prediction. SQUARNA seamlessly integrates sequence and alignment information with experimental data, such as residue reactivities obtained by chemical probing, as well as other structural restraints, including automated searches for Rfam database templates, G-quadruplex patterns, and protein-binding motifs. SQUARNA is available as a standalone tool at https://github.com/febos/SQUARNA and as a web server at https://larnal.imol.institute.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

1
Nucleic Acids Research
1281 papers in training set
Top 0.4%
30.0%
2
RNA
189 papers in training set
Top 0.2%
12.0%
3
Nature Communications
5641 papers in training set
Top 25%
6.5%
4
Bioinformatics
1204 papers in training set
Top 4%
4.7%
50% of probability mass above
5
NAR Genomics and Bioinformatics
242 papers in training set
Top 0.9%
4.7%
6
PLOS Computational Biology
1863 papers in training set
Top 8%
4.2%
7
Nature Methods
385 papers in training set
Top 3%
3.9%
8
Bioinformatics Advances
203 papers in training set
Top 2%
3.1%
9
RNA Biology
78 papers in training set
Top 0.4%
3.1%
10
Genome Biology
637 papers in training set
Top 5%
1.9%
11
Molecular Therapy Nucleic Acids
39 papers in training set
Top 0.6%
1.4%
12
PLOS ONE
5266 papers in training set
Top 52%
1.4%
13
Computational and Structural Biotechnology Journal
242 papers in training set
Top 5%
1.1%
14
Scientific Reports
3612 papers in training set
Top 68%
1.1%
15
Nature Biotechnology
172 papers in training set
Top 3%
1.1%
16
Journal of Molecular Biology
232 papers in training set
Top 3%
1.1%
17
Genomics, Proteomics & Bioinformatics
16 papers in training set
Top 0.1%
1.0%
18
Cell Reports Methods
165 papers in training set
Top 3%
1.0%
19
Molecular Biology and Evolution
542 papers in training set
Top 5%
1.0%
20
Biophysical Journal
631 papers in training set
Top 4%
1.0%
21
Genome Research
468 papers in training set
Top 7%
0.8%
22
Protein Science
246 papers in training set
Top 4%
0.8%
23
Journal of Chemical Theory and Computation
140 papers in training set
Top 1%
0.6%