Machine learning enables discovery of DNA-carbon nanotube sensors for serotonin
Kelich, P.; Jeong, S.; Navarro, N.; Adams, J.; Sun, X.; Zhao, H.; Landry, M. P.; Vukovic, L.
Show abstract
DNA-wrapped single walled carbon nanotube (SWNT) conjugates have remarkable optical properties leading to their use in biosensing and imaging applications. A critical limitation in the development of DNA-SWNT sensors is the current inability to predict unique DNA sequences that confer a strong analyte-specific optical response to these sensors. Here, near-infrared (nIR) fluorescence response datasets for ~100 DNA-SWNT conjugates, narrowed down by a selective evolution protocol starting from a pool of ~1010 unique DNA-SWNT candidates, are used to train machine learning (ML) models to predict new unique DNA sequences with strong optical response to neurotransmitter serotonin. First, classifier models based on convolutional neural networks (CNN) are trained on sequence features to classify DNA ligands as either high response or low response to serotonin. Second, support vector machine (SVM) regression models are trained to predict relative optical response values for DNA sequences. Finally, we demonstrate with validation experiments that integrating the predictions of ensembles of the highest quality CNN classifiers and SVM regression models leads to the best predictions of both high and low response sequences. With our ML approaches, we discovered five new DNA-SWNT sensors with higher fluorescence intensity response to serotonin than obtained previously. Overall, the explored ML approaches introduce an important new tool to predict useful DNA sequences, which can be used for discovery of new DNA-based sensors and nanobiotechnologies. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=100 SRC="FIGDIR/small/457145v1_ufig1.gif" ALT="Figure 1"> View larger version (37K): org.highwire.dtl.DTLVardef@9e1a2corg.highwire.dtl.DTLVardef@1c84e68org.highwire.dtl.DTLVardef@1939d1eorg.highwire.dtl.DTLVardef@30424e_HPS_FORMAT_FIGEXP M_FIG C_FIG
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Composite Hedges Nanopores: A High INDEL-Correcting Codec System for Rapid and Portable DNA Data Readout 95%
- De Novo Non-Canonical Nanopore Basecalling Enables Private Communication using Heavily-modified DNA Data at Single-Molecule Level 95%
- A digital twin for DNA data storage based on comprehensive quantification of errors and biases 95%
Similar papers in this journal
- Rapid Kinetic Fingerprinting of Single Nucleic Acid Molecules by a FRET-based Dynamic Nanosensor 94%
- Genetic Circuits Combined with Machine Learning Provides Fast Responding Living Sensors 93%
- Sensitive quantitative detection of SARS-CoV-2 in clinical samples using digital warm-start CRISPR assay 93%
Similar papers in this journal
Similar papers in this journal
- Enhanced Recognition of a Herbal Compound Epiberberine by a DNA Quadruplex-Duplex Structure 95%
- Flow-cell based technology for massively parallelcharacterization of base-modified DNA aptamers 94%
- Hairpin structure facilitates high-fidelity DNA amplification reactions in both qPCR and high-throughput sequencing 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.