choros: correction of sequence-based biases for accurate quantification of ribosome profiling data
Mok, A.; Tunney, R.; Benegas, G.; Wallace, E. W. J.; Lareau, L. F.
Show abstract
Ribosome profiling quantifies translation genome-wide by sequencing ribosome-protected fragments, or footprints. Its single-codon resolution allows identification of translation regulation, such as ribosome stalls or pauses, on individual genes. However, enzyme preferences during library preparation lead to pervasive sequence artifacts that obscure translation dynamics. Widespread over- and under-representation of ribosome footprints can dominate local footprint densities and skew estimates of elongation rates by up to five fold. To address these biases and uncover true patterns of translation, we present choros, a computational method that models ribosome footprint distributions to provide bias-corrected footprint counts. choros uses negative binomial regression to accurately estimate two sets of parameters: (i) biological contributions from codon-specific translation elongation rates; and (ii) technical contributions from nuclease digestion and ligation efficiencies. We use these parameter estimates to generate bias correction factors that eliminate sequence artifacts. Applying choros to multiple ribosome profiling datasets, we are able to accurately quantify and attenuate ligation biases to provide more faithful measurements of ribosome distribution. We show that a pattern interpreted as pervasive ribosome pausing near the beginning of coding regions is likely to arise from technical biases. Incorporating choros into standard analysis pipelines will improve biological discovery from measurements of translation.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- The genetic and biochemical determinants of mRNA degradation rates in mammals 95%
- Towards In-Silico CLIP-seq: Predicting Protein-RNA Interaction via Sequence-to-Signal Learning 95%
- DiMSum: an error model and pipeline for analyzing deep mutational scanning data and diagnosing common experimental pathologies 94%
Similar papers in this journal
- Correcting 4sU induced quantification bias in nucleotide conversion RNA-seq data 94%
- CorrAdjust unveils biologically relevant transcriptomic correlations by efficiently eliminating hidden confounders 94%
- High-resolution profiling reveals coupled transcriptional and translational regulation of transgenes 93%
Similar papers in this journal
- Accurate Estimation of Molecular Counts from Amplicon Sequence Data with Unique Molecular Identifiers 94%
- Ribo-ODDR: Ribo-seq focused Oligo Design pipeline for experiment-specific Depletion of Ribosomal RNAs 92%
- Defining data-driven primary transcript annotations with primaryTranscriptAnnotation in R 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.