Back

Phyllosphere and rhizosphere microbiomes empower Nicotiana tobacum complex traits dissection and prediction

Du, Q.; Han, Y.; Liu, Z.; Si, H.; Ji, Y.; Liu, L.; Xiao, Z.; Cheng, L.; Yang, A.; Liu, D.; Zan, Y.

2026-02-23 plant biology
10.64898/2026.02.22.707303 bioRxiv
Show abstract

Understanding how plant-associated microbiomes interact with host genome variation to influence agronomic traits is essential for advancing microbiomeassisted crop improvement. In this study, we characterized the phyllosphere and rhizosphere microbiomes of 164 diverse Nicotiana tabacum accessions using 16S rRNA sequencing and integrated these data with host genomic variation and 22 agronomic traits. The two microbiomes exhibited distinct taxonomic structures, diversity patterns, and predicted metabolic functions. Microbiome genomewide association studies identified extensive host genetic control over microbial abundance, including 49 shared genomic loci that explained nearly half of the heritable variation in both microbiomes. Microbiomewide association studies revealed biologically meaningful associations between specific ASVs and agronomic traits. However, network analysis demonstrated that microbial subcommunities, rather than individual taxa, contributed substantially to phenotypic variation. Then, colocalization analysis further identified genetic variants jointly influencing microbial abundance and metabolite traits, highlighting potential host-microbe-trait causal links. Incorporating microbiome data into genomic selection models, we successfully improved prediction accuracy for several traits, especially plant architecture and flowering. Together, this work provides a comprehensive populationlevel framework linking host genetics, microbiome composition, and agronomic traits in tobacco, offering new insights for microbiomeinformed breeding strategies.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.