Improving Bacterial Ribosome Profiling Data Quality
Glaub, A. S.; Huptas, C.; Neuhaus, K.; Ardern, Z.
Show abstract
Ribosome profiling (RIBO-seq) in prokaryotes has the potential to facilitate accurate detection of translation initiation sites, to increase understanding of translational dynamics, and has already allowed detection of many unannotated genes. However, protocols for ribosome profiling and corresponding data analysis are not yet standardized. To better understand the influencing factors, we analysed 48 ribosome profiling samples from 9 studies on E. coli K12 grown in LB medium. We particularly investigated the size selection step in each experiment since the selection for ribosome-protected footprints (RPFs) has been performed at various read lengths. We suggest choosing a size range between 22-30 nucleotides in order to obtain protein-coding fragments. In order to use RIBO-seq data for improving gene annotation of weakly expressed genes, the total amount of reads mapping to protein-coding sequences and not rRNA or tRNA is important, but no consensus about the appropriate sequencing depth has been reached. Again, this causes significant variation between studies. Our analysis suggests that 20 million non rRNA/tRNA mapping reads are required for global detection of translated annotated genes. Further, we highlight the influence of drug induced ribosome stalling, causing bias at translation start sites. Drug induced stalling may be especially useful for detecting weakly expressed genes. These suggestions should improve both gene detection and the comparability of resulting ribosome profiling datasets.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- From reporters to endogenous genes: the impact of the first five codons on translation efficiency in Escherichia coli 92%
- Identification of RNA 3' ends and termination sites in Haloferax volcanii 91%
- PresRAT: A server for identification of bacterial small-RNA sequences and their targets with probable binding region. 91%
Similar papers in this journal
- Efficient and cost-effective bacterial mRNA sequencing from low input samples through ribosomal RNA depletion 93%
- Complete genome sequence and annotation of the laboratory reference strain Shigella flexneri serovar 5a M90T and genome-wide transcriptional start site determination 93%
- Glucose-lactose mixture feeds in industry-like conditions: a gene regulatory network analysis on the hyperproducing Trichoderma reesei strain Rut-C30 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.