Comparative analysis of eccDNA and circRNA tools shows increased accuracy of tool combination
Zabala, A.; Ascension, A. M.; Prada-Luengo, I.; Otaegui, D.
Show abstract
IntroductionCircular nucleic acids such as extrachromosomal circular DNA (eccDNA) and circular RNA (circRNA) are increasingly recognized for their biological relevance and potential as biomarkers in disease contexts. Despite their growing importance, their detection remains challenging due to tool-specific biases, limited validation frameworks, and high variability in performance across datasets. MethodsWe benchmarked 10 circle detection tools across diverse conditions using both simulated and biological datasets. Our evaluation included classical performance metrics and a novel internal measure of read distribution symmetry ({Delta}CJ) to assess circle prediction confidence. We explored the impact of sequencing protocols, filtering strategies, and combined tool consensus. ResultsWe found that detection accuracy was highly influenced by sequencing depth, alignment algorithm, and experimental enrichment protocols. {Delta}CJ proved effective in flagging potential false positive circles, showing improved accuracy of Intersect (circles detected by all tools) and Rosette (circles detected by[≥] 2 tools) combinations. DiscussionThis study offers a broad evaluation of circular detection tools, suggesting that the combination of[≥] 3 tools is necessary for a correct prediction. These insights will inform future experimental design and data analysis pipelines in both experimental and clinical settings.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Tailored machine learning models for functional RNA detection in genome-wide screens 94%
- Improved characterization of single-cell RNA-seq libraries with paired-end avidity sequencing 94%
- Kmerator Suite: design of specific k-mer signatures andautomatic metadata discovery in large RNA-Seq datasets. 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.