Anthocyanin-associated cellular programs underlying terroir variation in Cabernet Sauvignon grape berry revealed by SEED-based deconvolution
Hu, X.; Tang, Y.; Deng, F.; Chen, Z.; Tang, G.; Yan, X.; Xia, Z.; Tong, H. H. Y.; Zhan, J.; Zou, X.; Hao, J.
Show abstract
Plant tissues consist of diverse cell populations that collectively contribute to development, metabolism, environmental responses, and phenotype formation. Although single-cell and single-nucleus RNA sequencing have greatly advanced the study of plant cellular heterogeneity, their application to large sample cohorts remains limited by cost, technical complexity, tissue dissociation constraints, and throughput. In contrast, bulk RNA-seq datasets have accumulated extensively across plant species, tissues, developmental stages, and environmental conditions, yet the celltype-level information embedded in these datasets remains difficult to resolve because plant-oriented deconvolution frameworks are still lacking. Existing deconvolution methods have largely been developed in mammalian systems and have not been systematically optimized for plant transcriptomic features, leaving their applicability under plant-specific constraints unclear. Here, we present SEED, an adaptive deconvolution framework optimized for plant transcriptomic data. SEED integrates candidate reference-template construction with seven deconvolution strategies and automatically identifies an optimal combination for a given dataset. In grapevine simulated benchmarking, SEED showed its clearest advantage under low-replication conditions and remained broadly competitive, rather than uniformly dominant, when larger pseudo-bulk sample sizes were evaluated. SEED further performed robustly in public Arabidopsis thaliana and Nicotiana tabacum datasets. Finally, we applied SEED to bulk RNA-seq data generated in this study from Vitis vinifera cv. Cabernet Sauvignon berries collected from Yinchuan and Yantai, identifying terroir-associated cell subtypes and coordinated celltype interaction patterns. Together, these results establish SEED as a practical framework for plant transcriptome deconvolution and provide a new tool for dissecting cellular heterogeneity associated with environmental adaptation and phenotype formation in plants.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genomic insights into longan evolution from a chromosome-level genome assembly and population genomics of longan accessions 93%
- Integrative analysis of the shikonin metabolic network identifies new gene connections and reveals evolutionary insight into shikonin biosynthesis 92%
- Telomere-to-telomere, gap-free genome of mung beans (Vigna radiata) provides insights into domestication under structural variation 92%
Similar papers in this journal
- Genome sequencing of 'Fuji' apple clonal varieties reveals evolutionary history and genetic mechanism of the spur-type morphology 93%
- A Day in the Life of Arabidopsis: 24-Hour Time-lapse Single-nucleus Transcriptomics Reveal Cell-type specific Circadian Rhythms 93%
- Reference-free assembly of long-read transcriptome sequencing data with RNA-Bloom2 93%
Similar papers in this journal
Similar papers in this journal
- Spatial and single-cell expression analyses reveal complex expression domains in early wheat spike development 93%
- Oatk: a de novo assembly tool for complex plant organelle genomes 93%
- Deciphering the Sequence Basis and Application of Transcriptional Initiation Regulation in Plant Genomes Through Deep Learning 93%
Similar papers in this journal
- Structural variations contribute to subspeciation and yield heterosis in rice 93%
- Chromosome-level Thlaspi arvense genome provides new tools for translational research and for a newly domesticated cash cover crop of the cooler climates 93%
- A chromosome-scale assembly of allotetraploid Brassica juncea (AABB) elucidates comparative architecture of the A and B genomes 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.