Generalizable direct protein sequencing with InstaNexus
Reverenna, M.; Wennekers Nielsen, M.; Wolff, D. S.; Lytra, E.; Colaianni, P. D.; Ljungars, A.; Laustsen, A. H.; Schoof, E. M.; Van Goey, J.; Jenkins, T. P.; Lukassen, M. V.; Santos, A.; Kalogeropoulos, K.
Show abstract
Protein-based therapeutics, such as antibodies and nanobodies, are not encoded in reference genomes, challenging their accurate characterization via standard proteomics. Current methods rely on indirect inference, fragmented outputs, and labor-intensive workflows, which hinder functional insights and routine application. Here, we present a generalizable, end-to-end workflow for direct protein sequencing, combining streamlined sample preparation, AI-driven de novo peptide sequencing, and tailored assembly to reconstruct contiguous protein sequences. A novel composite scoring framework prioritises longer assemblies and coverage, enhancing accuracy and reducing ambiguity. Validation across diverse protein modalities demonstrates its utility and ability to robustly sequence functionally critical regions of selected proteins. This workflow represents an advance in precision proteomics with promising applications in therapeutic discovery, immune profiling, and protein science.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- To fly, or not to fly, that is the question: A deep learning model for peptide detectability prediction in mass spectrometry 96%
- Mass spectrometry-based sequencing of the anti-FLAG-M2 antibody using multiple proteases and a dual fragmentation scheme 96%
- Automated Enrichment of Phosphotyrosine Peptides for High-Throughput Proteomics 96%
Similar papers in this journal
- Imputation of label-free quantitative mass spectrometry-based proteomics data using self-supervised deep learning 96%
- Systematic detection of functional proteoform groups from bottom-up proteomic datasets 95%
- SugarQuant: a streamlined pipeline for multiplexed quantitative site-specific N-glycoproteomics 95%
Similar papers in this journal
- Discriminating changes in protein structure using PTAD conjugation to tyrosine 94%
- Screening de novo designed protein binders in unpurified lysate using flow induced dispersion analysis 94%
- Phage display assisted discovery of a pH-dependent anti-alpha-cobratoxin antibody from a natural variable domain library 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.