Back

A meta-analysis of environmental sequencing data reveals the global distribution and hidden diversity of marine anaerobic ciliates

Schrecengost, A.; Frates, E.; Al-Haj, A. N.; Fulweiler, R. W.; Beinart, R. A.

2025-12-15 microbiology
10.64898/2025.12.15.694440 bioRxiv
Show abstract

Anaerobic protists are diverse, ecologically important members of anoxic microbial communities, acting as grazers, nutrient cyclers, and partners in multi-domain associations, yet remain understudied relative to anaerobic prokaryotes. Ciliates are particularly abundant and diverse in anoxia, but their global diversity and distribution are largely unknown. Here, we conducted a meta-analysis of public 18S rDNA datasets, along with one dataset generated here, to assess the global diversity and ecology of marine anaerobic ciliates. Using a novel pipeline, we processed 2854 samples from 42 studies spanning 19 habitat types. We recovered 3196 anaerobic ciliate amplicon sequence variants (ASVs) across all described lineages. Based on clade-specific divergence thresholds derived from phylogenetic distances, 28.4-46.3% of ASVs qualified as novel. Most sequences belonged to the poorly described plagiopylean family Epalxellidae, suggesting a large reservoir of undescribed diversity in this clade. Community comparisons revealed close phylogenetic similarities between some shallow-water and deep-sea assemblages, suggesting that shared redox conditions may shape communities more than water depth. Our results demonstrate that marine anaerobic ciliates are globally distributed, taxonomically diverse, and rich in novel lineages. This study provides a framework for leveraging environmental sequencing data to better understand the diversity and ecology of neglected protist lineages and under-sampled habitats.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.