Refining Cell-free DNA Metagenomics for Sepsis in Acute Leukemia: Towards Standardized Computational Strategies for Host Depletion and Microbial Detection from Shallow-Depth, Low-Coverage Profiles.
Mathur, A.; Anam, K.; Lall, S.; Paulose, E.; Terse, V.; Gawde, V.; Bhanshe, P.; Joshi, S.; Chaudhary, S.; More, M.; Salunke, P.; Tembhare, P.; Rajpal, S.; Chatterjee, G.; Subramanian, P. G.; Gujral, S.; Punatar, S.; Mirgh, S.; Chichra, A.; Bain, G.; Jindal, N.; Nayak, L.; Bagal, B.; Jain, H.; Sengar, M.; Shetty, A.; Gokarn, A.; Bhatt, V.; Khattry, N.; Patkar, N.
Show abstract
IntroductionAccurate alignment is critical for metagenomic profiling of plasma cell-free DNA (cfDNA), where microbial fragments are short, low in abundance, and often obscured by abundant human DNA. Yet, no unified guideline exists for selecting alignment algorithms in cfDNA studies. ResultsWe benchmarked six alignment methods (BWA-MEM2, Minimap2, Bowtie2-sensitive, Bowtie2-very-sensitive, Bowtie2-very-sensitive-local, and Kraken2) using in silico simulated datasets and cfDNA from 102 acute leukemia samples with sepsis. Evaluation criteria included bacterial read retention after human read filtering, minimization of false positives in real cfDNA datasets, and detection sensitivity for antimicrobial resistance (AMR) genes. BWA-MEM2 emerged as the most effective for removing stringent human reads and bacterial classification, while bowtie2(very_sensitive_local mode) was optimal for AMR gene detection. Post-filtering, cfDNA profiles exhibited marked heterogeneity in bacterial and AMR gene signatures, reflecting expected clinical variability. ConclusionsBWA-MEM2 provides optimal human read removal and bacterial taxa classification, minimizing false positives, though some true reads may be lost. For AMR gene detection, Bowtie2 (very_sensitive_local mode) demonstrated superior sensitivity in simulated datasets. This study presents the first systematic evaluation of read-based and k-mer-based aligners in plasma cfDNA, offering practical guidance for balancing human read removal with bacterial and AMR gene profiling to achieve robust and clinically meaningful results.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Functional profiling of COVID-19 respiratory tract microbiomes 94%
- Evaluation of microbiome enrichment and host DNA depletion in human vaginal samples using Oxford Nanopore's adaptive sequencing 94%
- MetaGut: Insights into gut microbiomes in stem cell transplantation by comprehensive shotgun long-read sequencing 93%
Similar papers in this journal
- Benchmarking taxonomic classifiers with Illumina and Nanopore sequence data for clinical metagenomic diagnostic applications 95%
- Metagenomic evidence for a polymicrobial signature of sepsis 95%
- Comparison of R9.4.1/Kit10 and R10/Kit12 Oxford Nanopore flowcells and chemistries in bacterial genome reconstruction 94%
Similar papers in this journal
- UltraSEQ: a universal bioinformatic platform for information-based clinical metagenomics and beyond 95%
- Machine-learning based detection of adventitious microbes in T-cell therapy cultures using long read sequencing 95%
- GenomicGapID: Leveraging Spatial Distribution of Conserved Genomic Sites for Broad-Spectrum Microbial Identification 94%
Similar papers in this journal
- Accurate and Reproducible Whole-Genome Genotyping for Bacterial Genomic Surveillance with Nanopore Sequencing Data 94%
- Clinical Metagenomic Sequencing for Species Identification and Antimicrobial Resistance Prediction in Orthopaedic Device Infection 94%
- Hash-based core genome multi-locus sequencing typing for Clostridium difficile 94%
Similar papers in this journal
- A comprehensive update to the Mycobacterium tuberculosis H37Rv reference genome 94%
- Convergence of resistance and evolutionary responses in Escherichia coli and Salmonella enterica co-inhabiting chicken farms in China 93%
- Olivar: automated variant aware primer design for multiplex tiled amplicon sequencing of pathogens 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.