G4All: a database of experimentally confirmed G-quadruplex-forming sequences
Parth, R.; Cucchiarini, A.; Trubetskoy, D.; Ferrari, G.; Chen, Y.; Baigum, M.; Guittat, L.; Lacroix, L.; Mergny, J.-L.
Show abstract
G-quadruplexes (G4s) are non-canonical nucleic acid structures with critical roles in gene regulation, genomic stability, and disease, making them prime targets for therapeutic and biotechnological applications. However, the absence of a curated, centralized, experimentally validated repository of short (mostly synthetic) G4-forming sequences, paired with appropriate single-stranded or hairpin controls, has limited reproducibility and hindered progress in the field. Here, we introduce G4All, a comprehensive, curated database of G4-forming short DNA and RNA sequences, complemented by rigorously selected non-G4 controls, all studied under roughly similar experimental conditions (around 100 mM potassium ion at near neutral pH). G4All consolidates sequences validated by diverse experimental methods, including circular dichroism, NMR, UV spectroscopy, with standardized annotations for topology, thermal stability, and experimental conditions. By providing both positive and negative datasets, G4All enables rigorous comparative analyses, assay benchmarking, and the development of predictive models. The database supports a broad range of applications, from fundamental studies of G4 biophysics to the rational design of aptamers, as well as benchmarking prediction algorithms. Future developments will expand G4All to include user-submitted datasets and more RNA sequences. Freely accessible, G4All offers a searchable interface and downloadable datasets, establishing a community-driven hub to accelerate discovery and standardization in G4 research. GRAPHICAL ABSTRACT O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=81 SRC="FIGDIR/small/743020v1_ufig1.gif" ALT="Figure 1"> View larger version (30K): org.highwire.dtl.DTLVardef@34c9f3org.highwire.dtl.DTLVardef@1b6a6feorg.highwire.dtl.DTLVardef@8d862borg.highwire.dtl.DTLVardef@1637155_HPS_FORMAT_FIGEXP M_FIG C_FIG
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A hybrid structure determination approach to investigate the druggability of the nucleocapsid protein of SARS-CoV-2 93%
- The Effect of Pseudoknot Base Pairing on Cotranscriptional Structural Switching of the Fluoride Riboswitch 92%
- Critical structural perturbations of ribozyme active sites induced by 2'-O-methylation commonly used in structural studies 92%
Similar papers in this journal
- Mechanistic Analysis of Riboswitch Ligand Interactions Provides Insights into Pharmacological Control over Gene Expression 95%
- Structure of the catalytically active APOBEC3G bound to a DNA oligonucleotide inhibitor reveals tetrahedral geometry of the transition state 93%
- Nanomechanics and co-transcriptional folding of Spinach and Mango 92%
Similar papers in this journal
- Expanded gene targeting in RNA hacking with G-tract-supply Staple oligomer 94%
- Sequence-Specific Installation of Aryl Groups in RNA via DNA-Catalyst Conjugates 92%
- Machine learning-augmented molecular dynamics simulations (MD) reveal insights into the disconnect between affinity and activation of ZTPriboswitch ligands 92%
Similar papers in this journal
- R-BIND 2.0: An Updated Database of Bioactive RNA-Targeting Small Molecules and Associated RNA Secondary Structures 94%
- Design of an orally bioavailable small molecule that modulates the microtubule-associated protein tau's pre-mRNA splicing 93%
- Bioorthogonal cyclopropenones for investigating RNA structure 92%
Similar papers in this journal
- ExploreTurns: A web tool for the exploration, analysis, and classification of beta turns and structured loops in proteins; application to beta-bulge and Schellman loops, Asx helix caps, beta hairpins and other hydrogen-bonded motifs 91%
- Identifying interactions between TDP-43 N-terminal and RNA binding domains 91%
- Arginine multivalency stabilizes protein/RNA condensates 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.