Improving N-Glycosylation and Biopharmaceutical Production Predictions Using AutoML-Built Residual Hybrid Models
Seber, P.; Braatz, R. D.
Show abstract
N-glycosylation has many essential biological roles, and is important for biotherapeutics as it can affect drug efficacy, duration of effect, and toxicity. Its importance has motivated the development of mechanistic models for quantitatively predicting the distribution of N-glycans during therapeutic protein production. Here we present a residual hybrid modeling approach that integrates mechanistic modeling with machine learning to produce significantly more accurate predictions for production of monoclonal antibodies in batch, fed-batch, and perfusion cell culture. For the largest dataset, the residual hybrid models have an average 736-fold reduction in testing prediction error. Furthermore, the residual hybrid models have lower prediction errors than the mechanistic models for all of the predicted variables in the datasets. We provide the automatic machine learning software used in this work, allowing other researchers to reproduce this work and use our software for other tasks and datasets.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Discriminating the Single-cell Gene Regulatory Networks of Human Pancreatic Islets: A Novel Deep Learning Application 92%
- AE-LGBM: Sequence-Based Novel Approach To Detect Interacting Protein Pairs via Ensemble of Autoencoder and LightGBM. 92%
- Employing Machine Learning Techniques to Detect Protein-Protein Interaction: A Survey, Experimental, and Comparative Evaluations 92%
Similar papers in this journal
- DeepInsight-3D for precision oncology: an improved anti-cancer drug response prediction from high-dimensional multi-omics data with convolutional neural networks 93%
- Designing of thermostable proteins with a desired melting temperature 93%
- A novel interpretable deep transfer learning combining diverse learnable parameters for improved T2D prediction based on single-cell gene regulatory networks 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.