Back

Structural Analysis of Prostate Cancer N-Glycans Using Graph-Based Structural Metrics

Kalyanthaya, M.; Gallegos, D.; Chow, J.; Pavletic, B.; Diaz Fernandez, A. B.; Kilcoyne, M.; Joshi, L.; Nguyen, D. H.

2026-06-14 bioinformatics
10.64898/2026.06.12.731995 bioRxiv
Show abstract

The N-linked glycans are structurally complex carbohydrate modifications that regulate protein folding, immune recognition, and cellular signaling, and their expression is extensively remodeled during cancer progression, making them promising biomarkers. In this study, prostate cancer-associated N-glycans from a range of relevant peer-reviewed studies were curated and digitized to develop a versatile computational framework that quantitatively encodes their spatial complexity across diverse biological systems. We invented two indices--the Distance & Connectivity Index (DCI) and the Position & Composition Index (PCI)--to capture the spatial information in N-glycans as layered architectures, enabling calculation of residue-level path lengths, branching structure, and compositional diversity. DCI summarizes glycan structure as both a scalar and matrix representation, while PCI does the same but also captures monosaccharide diversity, linkage heterogeneity, and cross-layer branching features. These metrics were computed with GlycoAssessor, an open-source platform that extracts information for the DCI and PCI from glycans drawn via Symbol Nomenclature for Glycans (SNFG) notation. Principal Component Analysis (PCA) was applied to evaluate whether glycans from prostate cancer tissues cluster distinctly in a disease-relevant manner. Results show that the spatial information in N-glycans: (1) increased in a multi-dimensional, non-linear manner, (2) objectively segregated structural themes, (3) could function as a potential prostate cancer biomarker that is distinct from mass-to-charge ratio and relative abundance, and (4) could objectively quantify novel subtype classifications of glycans associated with disease states and progression.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
Molecular & Cellular Proteomics
25 papers in training set
Top 0.1%
26.6%
2
Glycobiology
35 papers in training set
Top 0.1%
7.3%
3
Computational and Structural Biotechnology Journal
242 papers in training set
Top 0.3%
6.7%
4
Genomics, Proteomics & Bioinformatics
172 papers in training set
Top 0.4%
5.5%
5
Scientific Reports
3612 papers in training set
Top 18%
5.2%
50% of probability mass above
6
PLOS ONE
5266 papers in training set
Top 31%
4.9%
7
Journal of Proteome Research
234 papers in training set
Top 0.7%
4.0%
8
PLOS Computational Biology
1863 papers in training set
Top 9%
4.0%
9
Briefings in Bioinformatics
354 papers in training set
Top 3%
3.4%
10
Bioinformatics
1204 papers in training set
Top 6%
2.6%
11
PROTEOMICS
43 papers in training set
Top 0.3%
2.6%
12
Advanced Science
286 papers in training set
Top 3%
2.6%
13
Molecular & Cellular Proteomics
158 papers in training set
Top 0.8%
2.1%
14
Analytical Chemistry
218 papers in training set
Top 2%
1.5%
15
Communications Biology
993 papers in training set
Top 19%
1.3%
16
ACS Omega
105 papers in training set
Top 2%
1.1%
17
Journal of Chemical Information and Modeling
238 papers in training set
Top 2%
1.1%
18
Frontiers in Bioinformatics
49 papers in training set
Top 1%
1.0%
19
Frontiers in Chemistry
16 papers in training set
Top 0.5%
0.6%
20
GigaScience
212 papers in training set
Top 5%
0.6%
21
Computers in Biology and Medicine
128 papers in training set
Top 5%
0.6%
22
PNAS Nexus
159 papers in training set
Top 4%
0.6%
23
Frontiers in Endocrinology
58 papers in training set
Top 2%
0.6%
24
Frontiers in Cellular and Infection Microbiology
109 papers in training set
Top 4%
0.6%