Back

Insights from the genomes of four diploid Camelina spp.

Martin, S. L.; Toro, B. L.; James, T.; Sauder, C. A.; Laforest, M.

2021-08-23 genomics
10.1101/2021.08.23.455123 bioRxiv
Show abstract

Plant evolution has been a complex process involving hybridization and polyploidization making understanding the origin and evolution of a plants genome challenging even once a published genome is available. The oilseed crop, Camelina sativa (Brassicaceae), has a fully sequenced allohexaploid genome with three unknown ancestors. To better understand which extant species best represent the ancestral genomes that contributed to C. sativas formation, we sequenced and assembled chromosome level draft genomes for four diploid members of Camelina: C. neglecta C. hispida var. hispida, C. hispida var. grandiflora and C. laxa using long and short read data scaffolded with proximity data. We then conducted phylogenetic analyses on regions of synteny and on genes described for Arabidopsis thaliana, from across each nuclear genome and the chloroplasts to examine evolutionary relationships within Camelina and Camelineae. We conclude that C. neglecta is closely related to C. sativas sub-genome 1 and that C. hispida var. hispida and C. hispida var. grandiflora are most closely related to C. sativas sub-genome 3. Further, the abundance and density of transposable elements, specifically Helitrons, suggest that the progenitor genome that contributed C. sativas sub-genome 3 maybe more similar to the genome of C. hispida var. hispida than that of C. hispida var. grandiflora. These diploid genomes show few structural differences when compared to C. sativas genome indicating little change to chromosome structure following allopolyploidization. This work also indicates that C. neglecta and C. hispida are important resources for understanding the genetics of C. sativa and potential resources for crop improvement.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.