Exponentially few RNA structures are designable
Yao, H.-T.; Chauve, C.; REGNIER, M.; Ponty, Y.
Show abstract
The problem of RNA design attempts to construct RNA sequences that perform a predefined biological function, identified by several additional constraints. One of the foremost objective of RNA design is that the designed RNA sequence should adopt a predefined target secondary structure preferentially to any alternative structure, according to a given metrics and folding model. It was observed in several works that some secondary structures are undesignable, i.e. no RNA sequence can fold into the target structure while satisfying some criterion measuring how preferential this folding is compared to alternative conformations.\n\nIn this paper, we show that the proportion of designable secondary structures decreases exponentially with the size of the target secondary structure, for various popular combinations of energy models and design objectives. This exponential decay is, at least in part, due to the existence of undesignable motifs, which can be generically constructed, and jointly analyzed to yield asymptotic upper-bounds on the number of designable structures.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Algorithms to reconstruct past indels: the deletion-only parsimony problem 96%
- A regression based approach to phylogenetic reconstruction from multi-sample bulk DNA sequencing of tumors 95%
- Predicting Affinity Through Homology (PATH): Interpretable Binding Affinity Prediction with Persistent Homology 95%
Similar papers in this journal
- Maximum Mutational Robustness in Genotype-Phenotype Maps Follows a Self-similar Blancmange-like Curve 97%
- Efficient Manipulation and Generation of Kirchhoff Polynomials for the Analysis of Non-equilibrium Biochemical Reaction Networks 96%
- Predicting phenotype transition probabilities via conditional algorithmic probability approximations 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.