IsoSpace: Chemistry-Informed Dimensionality Reduction and Automated Isotope Candidate Detection in Imaging Mass Spectrometry
Meenakshi, M.; Migas, L. G.; Djambazova, K. V.; Dufresne, M.; Spraggins, J. M.; Van de Plas, R.
Show abstract
Imaging mass spectrometry (IMS) experiments simultaneously map thousands of ion species throughout tissue. Its high-dimensionality makes human interpretation difficult and dimensionality reduction (DR) methods are common to facilitate exploration. Traditional DR methods, such as principal component analysis (PCA), are general approaches, designed to work across application domains. They are usually unaware of, and unable to exploit correlations specific to a particular measurement type. Compression-focused, such blind methods often deliver physically impossible latent patterns or make feature combinations that confuse rather than aid human understanding. We introduce a novel chemistry-informed DR method, IsoSpace, with built-in awareness of mass spectrometry-relevant patterns. Although generalizable, IsoSpace substantiates chemistry-informed as sensitive to potential isotopic relation-ships. Like traditional DR methods, IsoSpace groups mass-over-charge (m/z)-features to deliver a low-dimensional representation of IMS data. Unlike traditional methods, IsoSpace ensures that retrieved latent patterns constitute potential isotopic families. IsoSpaces decomposition facilitates IMS interpretation at the molecular species rather than ion species level, implicitly automating isotope candidate detection. Most de-isotoping techniques ignore spatial relationships or rely on molecular class assumptions, making them less suitable for molecularly diverse tissue environments. IsoSpace avoids such assumptions, integrating spatial and spectral cues to empirically detect potential isotopic peaks. IsoSpace uses non-negative matrix factorization and an m/z -pattern matrix to uncover isotopic-like sequences, evaluating them by intra-pattern correlation. In a mouse pup example with 879 m/z -peaks, IsoSpace identified 71 potential isotopic patterns, substantially reducing data complexity while preserving chemical un-derstanding. IsoSpace o!ers chemical-interpretation-permissive DR with unsupervised isotope candidate detection for heterogeneous samples. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=83 SRC="FIGDIR/small/691393v1_ufig1.gif" ALT="Figure 1"> View larger version (24K): org.highwire.dtl.DTLVardef@5e8977org.highwire.dtl.DTLVardef@93150eorg.highwire.dtl.DTLVardef@4b7aa4org.highwire.dtl.DTLVardef@160aa43_HPS_FORMAT_FIGEXP M_FIG O_FLOATNOFigure 0:C_FLOATNO Graphical abstract for IsoSpace, chemistry-informed dimensionality reduction and au-tomated isotope candidate detection in imaging mass spectrometry. C_FIG
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Comparison of Cosine, Modified Cosine, and Neutral Loss Based Spectrum Alignment For Discovery of Structurally Related Molecules 96%
- MealTime-MS: A Machine Learning-Guided Real-Time Mass SpectrometryAnalysis for Protein Identification and Efficient DynamicExclusion 95%
- ModiFinder: Tandem Mass Spectral Alignment Enables Structural Modification Site Localization 95%
Similar papers in this journal
- Preserving Full Spectrum Information in ImagingMass Spectrometry Data Reduction 96%
- MAFFIN: Metabolomics Sample Normalization Using Maximal Density Fold Change with High-Quality Metabolic Features and Corrected Signal Intensities 96%
- MSModDetector: A Tool for Detecting Mass Shifts and Post-Translational Modifications in Individual Ion Mass Spectrometry Data 96%
Similar papers in this journal
- Scan-Centric, Frequency-Based Method for Characterizing Peaks from Direct Injection Fourier transform Mass Spectrometry Experiments 97%
- Robust Moiety Model Selection Using Mass Spectrometry Measured Isotopologues 97%
- WiPP: Workflow for improved Peak Picking for Gas Chromatography-Mass Spectrometry (GC-MS) data 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.