Back

Telomere-to-telomere Schizosaccharomyces japonicus genome assembly reveals hitherto unknown genome features

Etherington, G. J.; Nieduszynski, C.; Oliferenko, S.; Wu, P.-S.; Uhlmann, F.

2023-09-06 genomics
10.1101/2023.09.04.556195 bioRxiv
Show abstract

Schizosaccharomyces japonicus belongs to the single-genus class Schizosaccharomycetes, otherwise known as fission yeasts. As part of a composite model system with its widely studied S. pombe sister species, S. japonicus has provided critical insights into the workings and the evolution of cell biological mechanisms. Furthermore, its divergent biology makes S. japonicus a valuable model organism in its own right. However, the currently available short-read genome assembly contains gaps and has been unable to resolve centromeres and other repeat-rich chromosomal regions. Here we present a telomere-to-telomere long-read genome assembly of the S japonicus genome. This includes the three megabase-length chromosomes, with centromeres hundreds of kilobases long, rich in 5S ribosomal RNAs, transfer RNAs, long terminal repeats, and short repeats. We identify a gene-sparse region on chromosome 2 that resembles a 331 kb centromeric duplication. We revise the genome size of S. japonicus to at least 16.6 Mb and possibly up to 18.12 Mb, at least 30% larger than previous estimates. Our whole genome assembly will support the growing S. japonicus research community and facilitate research in new directions, including centromere and DNA repeat evolution, and yeast comparative genomics. Take-awayO_LIA telomere-to-telomere genome assembly of the fission yeast S. japonicus C_LIO_LIChromosome 2 harbours a previously unknown second centromere-like region C_LIO_LIThe estimated genome size of S. japonicus may be up to 18.12 Mb C_LI

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.