Back

Prediction of harvest-related traits in barley using high-throughput phenotyping data and machine learning

Tietze, H.; Abdelhakim, L.; Pleskacova, B.; Kurtz-Sohn, A.; Fridman, E.; Nikoloski, Z.; Panzarova, K.

2025-06-02 plant biology
10.1101/2025.05.29.656856 bioRxiv
Show abstract

Developing crop varieties that maintain productivity under drought is essential for future food security. Here, we investigated the potential of time-resolved high-throughput phenotyping to predict harvest-related traits and identify drought-stressed plants. Six barley lines (Hordeum vulgare) were grown in a greenhouse environment with well-watered and drought treatments, and phenotyped using RGB, thermal infrared, chlorophyll fluorescence and hyperspectral imaging sensors. Temporal phenomic classification model accurately distinguished between drought-treated and control plants, achieving high accuracy (R2 [≥] 0.97) even when exclusively using predictors only from the early phase after drought induction. Canopy temperature depression at the early stage and RGB-derived plant size estimates at the late stage were identified as key classification features. Temporal phenomic prediction model of harvest-related traits achieved particularly high mean R2 values for total biomass dry weight (0.97) and total spike weight (0.93), with RGB plant size estimators emerging as important predictors. Prediction accuracy for these traits remained high (R2 [≥] 0.84) when using only predictors from the first half of the experiment. Models trained on pooled drought and control data outperformed single-treatment models and retained high accuracy when applied across treatments. These findings support the integration of high-throughput phenotyping and temporal modelling to enable timely and more cost-effective selection of drought-resilient genotypes, and illustrate the broader potential of phenomics-driven approaches in accelerating crop improvement under stress-prone conditions.

Published in Frontiers in Plant Science (predicted rank #3) · training set

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.