Back

Challenges of Whole Genome Sequencing based Molecular Identification of Zoonotic Tuberculosis caused by Mycobacterium orygis

Soualhine, H.; Islam, R. M.; Sharma, M. K.; KhunKhun, R.; Shandro, C.; Sekirov, I.; Tyrell, G. J.

2023-02-28 microbiology
10.1101/2023.02.27.530375 bioRxiv
Show abstract

A recently described member of the Mycobacterium tuberculosis complex (MTBC) is Mycobacterium orygis, which can cause disease primarily in animals but also in humans. Although M. orygis has been reported from different geographic regions around the world, due to a lack of proper identification techniques, the contribution of this emerging pathogen to the global burden of zoonotic tuberculosis is not fully understood. In the present work, we report single nucleotide polymorphisms (SNPs) analysis using whole genome sequencing (WGS) that can accurately identify M. orygis and differentiate it from other members of MTBC species. WGS-based SNPs analysis was performed for 61 isolates from different Provinces in Canada that were identified as M. orygis. A total of 56 M. orygis sequences from the public databases were also included in the analysis. Several unique SNPs in gyrB, PPE55, Rv2042c, leuS, mmpL6, and mmpS6 genes were used to determine their effectiveness as genetic markers for the identification of M. orygis. To the best of our knowledge, five of these SNPs, viz., gyrB277 (A[->]G), gyrB1478 (T[->]C), leuS1064 (A[->]T), mmpL6486 (T[->]C), and mmpS6334 (C[->]G) are reported for the first time in this study. Our results also revealed several SNPs specific to other species within MTBC. The phylogenetic analysis shows that studied genomes were genetically diverse and clustered with M. orygis sequences of human and animal origin reported from different geographic locations. Therefore, the present study provides a new insight into the high confidence identification of M. orygis from MTBC species based on WGS data, which can be useful for reference and diagnostic laboratories.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.