Back

A Comprehensive Pangenome Approach to Exploring the Probiotic Potential of Weissella Confusa

Hashem, A.; Hossain, S.; Tuli, S. R.; Fatima, N.; Ali, S.; Moon, S. B.

2024-10-08 bioinformatics
10.1101/2024.10.06.616837 bioRxiv
Show abstract

BackgroundFermented foods harbour the bacterium Weissella confusa, which has probiotic properties but can potentially act as an opportunistic pathogen in humans and animals. Using pangenome analysis, our study aimed to improve the classification and identify functional traits of publicly available W. confusa genomes, focusing on evaluating their potential as probiotics. MethodThe genomic sequence of 120 strains of W. confusa was acquired from the NCBI RefSeq database. The downloaded sequences underwent a quality verification and filtering process. Ultimately, the chosen genomes were examined for comparative genomic analysis. We employed Roary to investigate the pangenome of W. confusa. Fishers exact test was utilised to analyse contingency tables, with a significance threshold of p < 0.05 (two-tailed). ResultsOur investigation revealed that the pangenome of W. confusa comprises 1100 core genes, 184 soft-core genes, 1407 shell genes, and 7006 cloud genes. This finding emphasises the "open" aspect of the W. confusa pangenome. The comparison of genomes showed that there were no acquired antibiotic resistance genes. However, the strains had different amounts of prophage regions, CRISPR arrays, and plasmids. Our research identified probiotic marker genes (PMGs), with the majority (78%) found in the core and soft-core genomes of various strains of W. confusa. ConclusionsAn extensive investigation of the W. confusa pangenome has concluded that it could be useful as a probiotic. Additional research is necessary to thoroughly evaluate the potential risks.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.