Back to BERT in 2026: ModernGENA as a Strong, Efficient Baseline for DNA Foundation Models
Aspidova, A.; Kuratov, Y.; Shadskiy, A.; Burtsev, M.; Fishman, V.
Show abstract
AO_SCPLOWBSTRACTC_SCPLOWRecent advances in DNA language models have mainly come from building larger and more complex architectures, making it harder to understand the effect of changes to standard components such as the transformer layers widely used in NLP. In this work, we study whether and how a modernized BERT-style back-bone (ModernBERT) can be adapted to genomic sequence modeling to improve computational efficiency, training stability, and downstream performance. Under controlled experimental settings, we benchmark efficiency across a range of sequence lengths and evaluate downstream performance on the Nucleotide Transformer benchmark. The resulting model, ModernGENA, achieves a strong efficiency-quality trade-off and ranks among the top-performing models in our evaluation suite. To support reproducibility and provide a solid default reference point for future architectural work in genomics, we release the full implementation and configuration of ModernGENA as an open, reusable baseline, and make ModernGENA base and ModernGENA large publicly available through the DNA language models collection on Hugging Face.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- EvoAug: improving generalization and interpretability of genomic deep neural networks with evolution-inspired data augmentations 95%
- Pair consensus decoding improves accuracy of neural network basecallers for nanopore sequencing 95%
- Enhancing sensitivity and controlling false discovery rate in somatic indel discovery 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.