Back

Statistical end-to-end analysis of large-scale microbial growth data with DGrowthR

Feldl, M.; Olayo-Alarcon, R.; Amstalden, M. K.; Zannoni, A.; Peschel, S.; Sharma, C. M.; Brochado, A. R.; Müller, C. L.

2025-03-25 microbiology
10.1101/2025.03.25.645164 bioRxiv
Show abstract

Quantitative analysis of microbial growth curves is essential for understanding how bacterial populations respond to environmental cues. Traditional analysis approaches make parametric assumptions about the functional form of these curves, limiting their usefulness for studying conditions that distort standard growth curves. In addition, modern robotics platforms enable the high-throughput collection of large volumes of growth data, thus requiring strategies that can analyze large-scale growth data in a flexible and efficient manner. Here, we introduce DGrowthR, a statistical R and standalone app frame-work for the integrative analysis of large growth experiments. DGrowthR comprises methods for data pre-processing and standardization, exploratory functional data analysis, and non-parametric modeling of growth curves using Gaussian Process regression. Importantly, DGrowthR includes a rigorous statistical testing framework for differential growth analysis. To illustrate the range of application scenarios of DGrowthR, we analyzed three large-scale bacterial growth datasets that tackle distinct scientific questions. On an in-house large-scale growth dataset comprising two pathogens that were subjected to a large chemical perturbation screen, DGrowthR enabled the discovery of compounds with significant growth inhibitory effects as well as compounds that induce non-canonical growth dynamics. We also re-analyzed two publicly available datasets and recovered reported adjuvants and antagonists of antibiotic activity, as well as bacterial genetic factors that determine susceptibility to specific antibiotic treatments. We anticipate that DGrowthR will streamline the analysis of modern high-volume growth experiments, enabling researchers to gain novel biological insights in a standardized and reproducible manner.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.