Back

Chromosome-scale Genome Assembly of Lewis Flax (Linum lewisii Pursh.)

Innes, P. A.; Smart, B. C.; Barham, J. A.; Hulke, B.; Kane, N. C.

2023-10-10 genomics
10.1101/2023.10.10.561607 bioRxiv
Show abstract

Linum lewisii, a perennial blue flax native to North America, holds potential as a sustainable perennial crop for oilseed production due to its ecological adaptability, upright harvestable structure, nutritious seeds, and low insect and disease issues. Its native distribution spans a large geographic range, from the Pacific Coast to the Mississippi River, and from Alaska to Baja California. Tolerant to cold and drought conditions, this species is also important for native ecosystem rehabilitation. Its enhancement of soil health, support for pollinators, and carbon sequestration underscore its agricultural relevance. This study presents a high-quality, chromosome-scale assembly of the L. lewisii (2n = 2x = 18) genome, derived from PacBio HiFi and Dovetail Omni-C sequencing of the "Maple Grove" variety. The initial assembly contained 642,903,787 base pairs across 2,924 scaffolds. Following HiRise scaffolding, the final assembly contained 643,041,835 base pairs, across 1,713 scaffolds, yielding an N50 contig length of 66,209,717 base pairs. Annotation of the assembly revealed 38,808 genes, including 37,599 protein-coding genes and 7,108 putative transposable elements. Analysis of synteny with other flax species revealed a striking number of chromosomal rearrangements. We also found an intriguing absence of the single-copy TSS1 gene in the L. lewisii genome, potentially linked to its transition from heterostyly to homostyly. Taken together, these findings represent a significant advancement in our understanding of the Linum genus and provide a resource for future domestication efforts and basic research on Lewis flax.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.