Reference-guided comparative genomics of seven Indonesian rice cultivars identifies conserved gene space and trait-associated sequence candidates
Purwestri, Y. A.; Wicaksono, A.; Nurbaiti, S.; Purba, N. T.; Retnaningati, D.; Restiani, R.; Kumalasari, N.; Nuringtyas, T. R.; Handayani, V. D. S.
Show abstract
Indonesian rice cultivars represent valuable genetic resources, yet many remain poorly characterized at the genomic level. Here, we generated 95.40 Gb of PacBio HiFi sequence data from seven Indonesian rice cultivars and constructed cultivar-specific consensus genomes using the telomere-to-telomere Nipponbare reference AGIS1.0. Sequencing coverage ranged from 27.92x to 41.58x, and the resulting consensus genomes spanned 387.93-390.54 Mb, with BUSCO completeness of approximately 98.3-98.5%. OrthoFinder assigned 99.1% of predicted proteins to 40,737 orthogroups, including 27,514 core orthogroups represented across all seven cultivars, indicating a highly conserved predicted gene space within the reference-guided framework. Targeted analysis recovered 278 of 280 cultivar-by-locus combinations representing 40 genes or gene family entries associated with grain pigmentation, nitrogen and amino-acid metabolism, and starch properties. Comparative predicted protein analysis prioritized ANS1, SBE2b, SSIIa/ALK, Wx/GBSSI, OsAAP6/qPC1, and SSI as candidates for further investigation. Among 269 completed AGIS1.0-anchored promoter comparisons, 159 passed quality-control criteria, whereas 110 were flagged for gene-model, boundary, synteny, or structural concerns. Notably, these flagged comparisons accounted for more than 90% of the alignment-derived sequence variation, emphasizing the importance of rigorous quality control when interpreting apparent promoter divergence. Collectively, these reference-guided genomic resources provide a standardized framework for investigating sequence variation in Indonesian rice germplasm and prioritize testable coding and regulatory candidates for functional validation and future genomics-assisted crop improvement.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A near gap-free haplotype-resolved genome assembly of Zoysia japonica uncovers intra-subgenomic gene expression and regulatory variation 94%
- Structural variations contribute to subspeciation and yield heterosis in rice 94%
- CRISPR-based editing of the ω- and γ-gliadin gene clusters reduces wheat immunoreactivity without affecting grain protein quality 93%
Similar papers in this journal
- Ubiquitin-like SUMO protease expansion in rice (Oryza sativa) 94%
- Genome-wide Population Structure Analyses of Three Minor Millets: Kodo Millet, Little Millet, and Proso Millet 93%
- Genome-wide association study of multiple yield components in a diversity panel of polyploid sugarcane (Saccharum spp.) 93%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Natural variation in expression of the HECT E3 ligase UPL3 influences seed size and crop yields in Brassica napus by altering regulatory gene expression. 92%
- Increased and ectopic expression of Triticum polonicum VRT-A2 underlies elongated glumes and grains in hexaploid wheat in a dosage-dependent manner 92%
- Functional Insights into Dispensable Genes using Genome-Wide Loss-of-Function Burden Tests in Arabidopsis 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.