Single-kernel near-infrared spectroscopy enables haploid kernel sorting in field and sweet corn using high-oil haploid inducers across diverse donor-inducer combinations
Sharma, S.; Gustin, J. L.; Frei, U. K.; Settles, A. M.; Lübberstedt, T.; Resende, M. F. R.; Hershberger, J.
Show abstract
Key messageA single-kernel near-infrared reflectance spectroscopy-based sorter can effectively identify haploid kernels for doubled haploid production in field and sweet corn backgrounds. Doubled haploid (DH) technology significantly shortens the breeding cycle for developing homozygous inbred lines in maize (Zea mays). Manual sorting of haploids from a larger bulk of hybrid kernels in an induction cross is a major bottleneck in DH development. Automated systems based on near-infrared (NIR) reflectance spectroscopy can be valuable tools for rapid haploid sorting, provided that sorting accuracy is sufficient for incorporation into the DH process. In this study, we evaluated the accuracy of a custom-built single-kernel NIR (skNIR) sorter for classifying haploid kernels from 12 high-oil haploid induction populations generated from two sweet corn and two field corn donors and four high-oil haploid inducers (HOHIs). We evaluated several general classification models that can be applied without population-specific recalibration or prior genotyping, including models that classified haploids based solely on predicted oil content, as well as multivariate methods that used all wavelengths of the NIR spectra. The highest classification accuracy was obtained using a general multivariate support vector machine (SVM) model. When combined with the two best-performing HOHIs, the general SVM model accurately sorted induction populations from two of the three donor backgrounds crossed with these inducers. Two oil-based methods showed less accurate classification than the multivariate SVM model, due to overlapping oil content distributions across the two kernel classes. Overall, this study demonstrates effective skNIR-based sorting of haploid kernels from diverse induction populations using a single general model. The practical deployment of this instrument in maize breeding programs is discussed.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Predicting Moisture Content During Maize Nixtamalization Using Machine Learning with NIR Spectroscopy 95%
- Predicted genetic gains from introgressing chromosome segments from exotic germplasm into an elite soybean cultivar 94%
- Multi-omics prediction of oat agronomic and seed nutritional traits across environments and in distantly related populations 94%
Similar papers in this journal
Similar papers in this journal
- Genetic diversity of North American popcorn germplasm and the effect of population structure on nicosulfuron response 94%
- Reciprocal recurrent selection based on genetic complementation: An efficient way to build heterosis in diploids due to directional dominance 94%
- Optimizing population simulations to accurately parallel empirical data for digital breeding 94%
Similar papers in this journal
- Deciphering the genetic diversity of landraces with high-throughput SNP genotyping of DNA bulks: methodology and application to the maize 50k array 94%
- PATRIOT: A pipeline for tracing identical-by-descent chromosome segments to improve genomic prediction in self-pollinating crop species 94%
- Comparison of Single-Trait and Multi-Trait Genome-Wide Association Models and Inclusion of Correlated Traits in the Dissection of the Genetic Architecture of a Complex Trait in a Breeding Program 94%
Similar papers in this journal
- Investigating genomic prediction strategies for grain carotenoid traits in a tropical/subtropical maize panel 96%
- Evaluating metabolic and genomic data for predicting grain traits under high night temperature stress in rice 94%
- Interchromosomal Linkage Disequilibrium Analysis Reveals Strong Indications of Sign Epistasis in Wheat Breeding Families 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.