A generalizable method for false-discovery rate estimation in mass spectrometry-based lipidomics
Fujimoto, G. M.; Kyle, J. E.; Lee, J.-Y.; Metz, T. O.; Payne, S. H.
Show abstract
Mass spectrometry (MS)-based lipidomics is revolutionizing lipid research with high throughput identification and quantification of hundreds to thousands of lipids with the goal of elucidating lipid metabolism and function. Estimates of statistical confidence in lipid identification are essential for downstream data interpretation in a biological context. In the related field of proteomics, a variety of methods for estimating false-discovery are available, and understanding the statistical confidence of identifications is typically required for data analysis and hypothesis testing. However, there is no current method for estimating the false discovery rate (FDR) or statistical confidence for MS-based lipid identifications. This has slowed the adoption of MS-based lipidomics research, as all identifications require manual inspection and validation to ensure their accuracy. We present here the first generalizable method for FDR estimation, a target/decoy approach, that allows those conducting MS-based lipidomics research to confidently adjust spectral score thresholds to minimize false discovery and to enable full automation of data analysis.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- MealTime-MS: A Machine Learning-Guided Real-Time Mass SpectrometryAnalysis for Protein Identification and Efficient DynamicExclusion 98%
- Comparison of Cosine, Modified Cosine, and Neutral Loss Based Spectrum Alignment For Discovery of Structurally Related Molecules 96%
- IS-PRM-based peptide targeting informed by long-read sequencing for alternative proteome detection 96%
Similar papers in this journal
- ReCom: A semi-supervised approach to ultra-tolerant database search for improved identification of modified peptides 97%
- AA_stat: intelligent profiling of in vivo and in vitro modifications from open search results 96%
- ibaqpy: A scalable Python package for baseline quantification in proteomics leveraging SDRF metadata 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.