Automated high-throughput biological sex identification from archaeological human dental enamel using targeted proteomics
Koenig, C.; Bortel, P.; Paterson, R. S.; Rendl, B.; Madupe, P. P.; Troche, G. B.; Hermann, N. V.; Martinez de Pinillos, M.; Martinon-Torres, M.; Mularczyk, S.; Jorkov, M. L. S.; Gerner, C.; Kanz, F.; Martinez-Val, A.; Cappellini, E.; Olsen, J. V.
Show abstract
Biological sex is key information for archaeological and forensic studies, which can be determined by proteomics. However, lack of a standardised approach for fast and accurate sex identification currently limits the reach of proteomics applications. Here, we introduce a streamlined mass spectrometry (MS)-based workflow for determination of biological sex using human dental enamel. Our approach builds on a minimally invasive sampling strategy by acid etching, a rapid online liquid chromatography (LC) gradient coupled to high-resolution parallel reaction monitoring assay allowing for a throughput of 200 samples-per-day with high quantitative performance enabling confident identification of both males and females. Additionally, we have developed a streamlined data analysis pipeline and integrated it into an R-Shiny interface for ease-of-use. The method was first developed and optimised using modern teeth and then validated in an independent set of deciduous teeth of known sex. Finally, the assay was successfully applied to archaeological material, enabling the analysis of over 300 individuals. We demonstrate unprecedented performance and scalability, speeding up MS analysis by tenfold compared to conventional proteomics-based sex identification methods. This work paves the way for large-scale archaeological or forensic studies enabling the investigation of entire populations rather than focusing on individual high-profile specimens.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Data-Driven Optimization of DIA Mass Spectrometry by DO-MS 97%
- Low-invasive sampling method for taxonomic for the identification of archaeological and paleontological bones by proteomics of their collagens 96%
- PAMPA: a software for peptide markers and taxonomic identification for ZooMS samples in Archaeology and Paleontology 96%
Similar papers in this journal
- Heat n Beat: A universal high-throughput end-to-end proteomics sample processing platform in under an hour 97%
- Optimization of data-independent acquisition using predicted libraries for deep and accurate proteome profiling 97%
- Ultra-sensitive nanoLC-MS of sub nanogram protein samples using second generation micro pillar array LC technology with Orbitrap Exploris 480 and FAIMS PRO 96%
Similar papers in this journal
- A combined flow injection/reversed phase chromatography - high resolution mass spectrometry workflow for accurate absolute lipid quantification with 13C- internal standards 95%
- Profiling embryonic stem cell differentiation by MALDI TOF mass spectrometry: development of a reproducible and robust sample preparation workflow 95%
- In-situ lipid profiling of insect pheromone glands by Direct Analysis in Real Time Mass Spectrometry 95%
Similar papers in this journal
- Standardization and Harmonization of Distributed Multi-National Proteotype Analysis supporting Precision Medicine Studies 96%
- OzFAD: Ozone-enabled fatty acid discovery reveals unexpected diversity in the human lipidome 96%
- Single Cell Proteomics Using a Trapped Ion Mobility Time-of-Flight Mass Spectrometer Provides Insight into the Post-translational Modification Landscape of Individual Human Cells 95%
Similar papers in this journal
- Development of a PNGase Rc column for online deglycosylation of complex glycoproteins during HDX-MS 96%
- Ultraviolet Photodissociation of tryptic peptide backbones at 213 nm 96%
- MealTime-MS: A Machine Learning-Guided Real-Time Mass SpectrometryAnalysis for Protein Identification and Efficient DynamicExclusion 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.