Back

Genome-wide mapping of the Galleria mellonella larvae transcription start sites during fungal infection and treatment

Abugessaisa, I.; Konings, M.; Manabe, R.-i.; Tagami, M.; Severin, J.; Hasegawa, A.; Kawashima, T.; Kinoshita, H.; Noma, S.; Takahashi, C.; Verbon, A.; Okazaki, Y.; van de Sande, W. W. J.; Kasukawa, T.

2025-01-22 genomics
10.1101/2025.01.20.627872 bioRxiv
Show abstract

Using Low Quantity single strand CAGE (LQ-ssCAGE), we mapped the transcription start sites (TSS). We annotated the 5 end of the invertebrate Galleria mellonella, an upcoming and booming experimental model in infectious disease and immunology research. However, the current genome annotation of this model organism lacks annotation of the 5 end and TSS information. G. mellonella larva was infected with the fungal pathogen Madurella mycetomatis to map TSS under healthy and infection conditions. Larvae were first treated with itraconazole or ravuconazole, and then RNA-seq and LQ-ssCAGE libraries were prepared and sequenced 4, 30, and 52 hours following infection. The LQ-ssCAGE data was processed to identify CAGE transcription start site (CTSS), uni-, and bi-directional clusters. LQ-ssCAGE enabled us to precisely identify 39,410 TSS and 249 active enhancers; we assigned genomic features to the resulting TSSs and enhancers. The majority of the TSS peaks are annotated as promoter regions, while the enhancers were annotated as intergenic and genic. Furthermore, we confirmed the quality of TSS calling by promoter shapes and GC bias. Furthermore, we identified a set of super-enhancers and predicted de-novo motifs. The raw and processed data was deposited to NCBI GEO GSE282923. CTSS, TSS peaks, and enhancers coordinated are available through the ZENBU Genome browser. In this study, we reported the first atlas of TSS and active enhancers of G. mellonella.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.