Back

A Diploid Panax Genome Reveals Ginsenoside Diversity Driven by UGT Diversification and Network Rewiring, Rather Than Gene Family Expansion

Xu, Z.; Li, W.; Wei, F.-g.; Xiong, G.; Chen, Z.-j.; Gao, L.-z.

2026-06-21 genomics
10.64898/2026.06.16.732563 bioRxiv
Show abstract

The medicinal herb Panax notoginseng produces a structurally diverse array of triterpene saponins (ginsenosides), yet the genetic basis of this chemical complexity remains unclear. Here we present a high-quality chromosome-level genome of diploid P. notoginseng and integrate comparative genomics with multi-tissue, multi-year metabolomics and transcriptomics. Surprisingly, unlike tetraploid Panax species, P. notoginseng shows no general expansion of core saponin biosynthetic gene families. Instead, lineage-specific diversification of UDP-glycosyltransferase (UGT) families, a recent burst of LTR retrotransposons, and enrichment of species-specific genes in metabolic modification pathways point to an alternative evolutionary route. Saponin accumulation follows strict spatiotemporal compartmentalisation, and co-expression network analysis reveals that the biosynthetic machinery is not static but continuously rewired during development-from a basic synthesis module in the first year to a modular pattern supporting both broad accumulation and branch-specific modification by the third year. Seventeen differentially expressed UGTs show clear tissue preferences and saponin-branch correlations. As a representative example, PnUGT33 is tightly linked to the PPD-type saponin branch; structural modelling, molecular docking and 100 ns molecular dynamics simulations demonstrate its differential recognition of diverse triterpene skeletons. Collectively, our findings establish that ginsenoside diversity in diploid P. notoginseng arises primarily from UGT lineage diversification, developmentally rewired regulatory networks and UGT mediated branch selective post-modification, rather than from expansion of core pathway genes. This work provides a new paradigm for understanding how plants achieve metabolic complexity without whole genome duplication or massive gene amplification.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Nature Plants
94 papers in training set
Top 0.1%
18.5%
2
Nature Communications
5641 papers in training set
Top 11%
15.1%
3
The Plant Cell
161 papers in training set
Top 0.7%
6.3%
4
Molecular Plant
39 papers in training set
Top 0.2%
5.2%
5
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 12%
4.4%
6
eLife
5828 papers in training set
Top 29%
4.0%
50% of probability mass above
7
Nature
645 papers in training set
Top 4%
3.5%
8
Science
477 papers in training set
Top 2%
3.5%
9
Nature Structural & Molecular Biology
18 papers in training set
Top 0.1%
3.4%
10
Nature Genetics
286 papers in training set
Top 2%
3.4%
11
Cell Reports
1498 papers in training set
Top 12%
3.3%
12
Cell
431 papers in training set
Top 3%
3.3%
13
New Phytologist
346 papers in training set
Top 3%
1.9%
14
Molecular Biology and Evolution
542 papers in training set
Top 3%
1.7%
15
Plant Communications
36 papers in training set
Top 0.5%
1.5%
16
Current Biology
665 papers in training set
Top 7%
1.4%
17
Communications Biology
993 papers in training set
Top 19%
1.3%
18
Nature Chemical Biology
119 papers in training set
Top 2%
1.1%
19
Science Advances
1243 papers in training set
Top 27%
1.0%
20
Scientific Reports
3612 papers in training set
Top 69%
1.0%
21
Cell Genomics
172 papers in training set
Top 4%
1.0%
22
Nucleic Acids Research
1281 papers in training set
Top 13%
0.8%
23
Cell Chemical Biology
94 papers in training set
Top 2%
0.8%
24
PLOS Biology
486 papers in training set
Top 12%
0.8%
25
Journal of Experimental Botany
219 papers in training set
Top 3%
0.6%
26
Plant Biotechnology Journal
64 papers in training set
Top 1%
0.6%
27
Genome Biology
637 papers in training set
Top 9%
0.6%
28
Horticulture Research
47 papers in training set
Top 1%
0.6%
29
iScience
1154 papers in training set
Top 39%
0.6%