Computational design and evaluation of optimal bait sets for scalable proximity proteomics
Kasmaeifar, V.; Sedighi, S.; Gingras, A.-C.; Campbell, K. R.
Show abstract
The spatial organization of proteins in eukaryotic cells can be explored by identifying nearby proteins using proximity-dependent biotinylation approaches like BioID. BioID defines the localization of thousands of endogenous proteins in human cells when used on hundreds of bait proteins. However, this high bait number restricts the approachs usage and gives these datasets limited scalability for context-dependent spatial profiling. To make subcellular proteome mapping across different cell types and conditions more practical and cost-effective, we developed a comprehensive benchmarking platform and multiple metrics to assess how well a given bait subset can reproduce an original BioID dataset. We also introduce GENBAIT, which uses a genetic algorithm to optimize bait subset selection, to derive bait subsets predicted to retain the structure and coverage of two large BioID datasets using less than a third of the original baits. This flexible solution is poised to improve the intelligent selection of baits for contextual studies.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A versatile information retrieval framework for evaluating profile strength and similarity 95%
- LEOPARD: missing view completion for multi-timepoints omics data via representation disentanglement and temporal knowledge transfer 94%
- Scarf: A toolkit for memory efficient analysis of large-scale single-cell genomics data 94%
Similar papers in this journal
- Protein prediction models support widespread post-transcriptional regulation of protein abundance by interacting partners 94%
- PathIntegrate: Multivariate modelling approaches for pathway-based multi-omics data integration 94%
- Validation and tuning of in situ transcriptomics image processing workflows with crowdsourced annotations 93%
Similar papers in this journal
- Cracking the black box of deep sequence-based protein-protein interaction prediction 94%
- Supervised Application of Internal Validation Measures to Benchmark Dimensionality Reduction Methods in scRNA-seq Data 93%
- Assessing deep learning algorithms in cis-regulatory motif finding based on genomic sequencing data 93%
Similar papers in this journal
Similar papers in this journal
- EPIFANY - A method for efficient high-confidence protein inference 94%
- Pairwise Attention: Leveraging Mass Differences to Enhance De Novo Sequencing of Mass Spectra 94%
- Biological Function Assignment Across Taxonomic Levels in Mass-Spectrometry-Based Metaproteomics via a Modified Expectation Maximization Algorithm 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.