Back

Telomere-to-telomere African wild rice (Oryza longistaminata) reference genome reveals segmental and structural variation

Guang, X.; Yang, J.; Zhang, S.; Guo, F.; Li, L.; Lian, X.; Zeng, T.; Cai, C.; Liu, F.; Li, Z.; Hu, Y.; Fang, D.; He, W.; Sahu, S. K.; Li, W.; Lu, H.; Li, Y.; Liu, H.; Xu, X.; Guang, Y.; Hu, F.; Dong, Y.; Wei, T.

2024-09-07 genomics
10.1101/2024.09.05.611405 bioRxiv
Show abstract

Rice (Oryza sativa) is one of the most important staple food crops worldwide, and its wild relatives serve as an important gene pool in its breeding. Compared with cultivated rice species, African wild rice (Oryza longistaminata) has several advantageous traits, such as resistance to increased biomass production, clonal propagation via rhizomes, and biotic stresses. However, previous O. longistaminata genome assemblies have been hampered by gaps and incompleteness, restricting detailed investigations into their genomes. To streamline breeding endeavors and facilitate functional genomics studies, we generated a 343-Mb telomere-to-telomere (T2T) genome assembly for this species, covering all telomeres and centromeres across the 12 chromosomes. This newly assembled genome has markedly improved over previous versions. Comparative analysis revealed a high degree of synteny with previously published genomes. A large number of structural variations were identified between the O. longistaminata and O. sativa. A total of 2,466 segmentally duplicated genes were identified and enriched in cellular amino acid metabolic processes. We detected a slight expansion of some subfamilies of resistance genes and transcription factors. This newly assembled T2T genome of O. longistaminata provides a valuable resource for the exploration and exploitation of beneficial alleles present in wild relative species of cultivated rice.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.