Back

Characterising Aromatic Side Chains in Proteins through the Synergistic Development of NMR Experiments and Deep Neural Networks

Shukla, V. K.; Karunanithy, G.; Vallurupalli, P.; Hansen, D. F.

2024-04-02 biophysics
10.1101/2024.04.01.587635 bioRxiv
Show abstract

Nuclear magnetic resonance (NMR) spectroscopy has become an important technique in structural biology for characterising the structure, dynamics and interactions of macromolecules. While a plethora of NMR methods are now available to inform on backbone and methyl-bearing side-chains of proteins, a characterisation of aromatic side chains is more challenging and often requires specific labelling or 13C-detection. Here we present a deep neural network (DNN) named FID-Net-2, which transforms NMR spectra recorded on simple uniformly 13C labelled samples to yield high-quality 1H-13C correlation spectra of the aromatic side chains. Key to the success of the DNN is the design of a complementary set of NMR experiments that produce spectra with unique features to aid the DNN produce high-resolution aromatic 1H-13C correlation spectra with accurate intensities. The reconstructed spectra can be used for quantitative purposes as FID-Net-2 predicts uncertainties in the resulting spectra. We have validated the new methodology experimentally on protein samples ranging from 7 to 40 kDa in size. We demonstrate that the method can accurately reconstruct high resolution two-dimensional aromatic 1H-13C correlation maps, high resolution three-dimensional aromatic-methyl NOESY spectra to facilitate aromatic 1H-13C assignments, and that the intensities of peaks from the reconstructed aromatic 1H-13C correlation maps can be used to quantitatively characterise the kinetics of protein folding. More generally, we believe that this strategy of devising new NMR experiments specifically for analysis using customised DNNs represents a substantial advance that will have a major impact on the study of molecules using NMR in the years to come.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.