A computational strategy to uncover fusion genes in prostate cancer cell lines
Morgan, R. A.; Hardiman, G.
Show abstract
Fusion genes, chimeric transcripts formed by the merging of two distinct genes due to chromosomal structural changes (e.g., inversions or trans/cis-splicing), are established cancer drivers. Advances in genomic technologies, particularly RNA sequencing and improved fusion gene prediction algorithms, have significantly expanded our understanding of fusion genes in cancer. This chapter explores computational methods for identifying fusion genes through RNA sequencing data, using the TMPRSS2::ERG fusion in prostate cancer cell lines as a case study, and includes analysis of both fusion-positive and fusion-negative cell lines. To achieve high-confidence detection, three open-source fusion prediction tools, STAR-Fusion, FusionCatcher, and JAFFA are investigated. These tools were selected for their accessibility, active maintenance, and strong performance in benchmarking studies. Their sensitivity and accuracy in detecting TMPRSS2::ERG is systematically evaluated and validated, ensuring robust and high-resolution detection of fusion events.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Comparative analysis of ChIP-exo peak-callers: impact of data quality, read duplication and binding subtypes 95%
- scMuffin: an R package for disentangling solid tumor heterogeneity from single-cell expression data 94%
- ILIAD: A suite of automated Snakemake workflows for processing genomic data for downstream applications 94%
Similar papers in this journal
- iCOMIC: a graphical interface-driven bioinformatics pipeline for analyzing cancer omics data 96%
- FLYNC: A Machine Learning-Driven Framework for Discovering Long Non-Coding RNAs in Drosophila melanogaster 96%
- Dashboard-style interactive plots for RNA-seq analysis are R Markdown ready with Glimma 2.0 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.