Back

bletl - A Python package for integrating microbioreactors in the design-build-test-learn cycle

Osthege, M.; Tenhaef, N.; Zyla, R.; Mueller, C.; Hemmerich, J.; Wiechert, W.; Noack, S.; Oldiges, M.

2021-08-25 bioengineering
10.1101/2021.08.24.457462 bioRxiv
Show abstract

Microbioreactor (MBR) devices have emerged as powerful cultivation tools for tasks of microbial phenotyping and bioprocess characterization and provide a wealth of online process data in a highly parallelized manner. Such datasets are difficult to interpret in short time by manual workflows. In this study, we present the Python package bletl and show how it enables robust data analyses and the application of machine learning techniques without tedious data parsing and preprocessing. bletl reads raw result files from BioLector I, II and Pro devices to make all the contained information available to Python-based data analysis workflows. Together with standard tooling from the Python scientific computing ecosystem, interactive visualizations and spline-based derivative calculations can be performed. Additionally, we present a new method for unbiased quantification of time-variable specific growth rate [Formula] based on a novel method of unsupervised switchpoint detection with Student-t distributed random walks. With an adequate calibration model, this method enables practitioners to quantify time-variable growth rate with Bayesian uncertainty quantification and automatically detect switch-points that indicate relevant metabolic changes. Finally, we show how time series feature extraction enables the application of machine learning methods to MBR data, resulting in unsupervised phenotype characterization. As an example, t-distributed Stochastic Neighbor Embedding (t-SNE) is performed to visualize datasets comprising a variety of growth/DO/pH phenotypes. Practical ApplicationThe bletl package can be used to analyze microbioreactor datasets in both data analysis and autonomous experimentation workflows. Using the example of BioLector datasets, we show that loading such datasets into commonly used data structures with one line of Python code is a significant improvement over spreadsheet or hand-crafted scripting approaches. On top of established standard data structures, practitioners may continue with their favorite data analysis routines, or make use of the additional analysis functions that we specifically tailored to the analysis of microbioreactor time series. Particularly our function to fit cross-validated smoothing splines can be used for on-line signals from any microbioreactor system and has the potential to improve robustness and objectivity of many data analyses. Likewise, our random walk based [Formula] method for inferring growth rates under uncertainty, but also the time-series feature extraction may be applied to on-line data from other cultivation systems as well.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Journal of Open Source Software
25 papers in training set
Top 0.1%
15.6%
2
PLOS Computational Biology
1863 papers in training set
Top 2%
15.6%
3
PLOS ONE
5266 papers in training set
Top 21%
8.2%
4
SoftwareX
15 papers in training set
Top 0.1%
5.7%
5
Computer Methods and Programs in Biomedicine
28 papers in training set
Top 0.2%
3.6%
6
Ecological Informatics
33 papers in training set
Top 0.2%
3.3%
50% of probability mass above
7
Plant Phenomics
18 papers in training set
Top 0.1%
3.3%
8
HardwareX
18 papers in training set
Top 0.1%
2.7%
9
Bioinformatics
1204 papers in training set
Top 6%
2.7%
10
IFAC-PapersOnLine
13 papers in training set
Top 0.1%
2.2%
11
Scientific Reports
3612 papers in training set
Top 50%
2.0%
12
Nature Communications
5641 papers in training set
Top 44%
1.8%
13
Computational and Structural Biotechnology Journal
242 papers in training set
Top 4%
1.6%
14
Methods in Ecology and Evolution
176 papers in training set
Top 1%
1.5%
15
GigaScience
212 papers in training set
Top 3%
1.4%
16
Frontiers in Bioengineering and Biotechnology
98 papers in training set
Top 2%
1.2%
17
eLife
5828 papers in training set
Top 55%
1.2%
18
NAR Genomics and Bioinformatics
242 papers in training set
Top 4%
1.1%
19
Frontiers in Bioinformatics
49 papers in training set
Top 1.0%
1.1%
20
Methods
34 papers in training set
Top 0.5%
1.1%
21
Philosophical Transactions of the Royal Society A: Mathematical, Physical and Engineering Sciences
12 papers in training set
Top 0.1%
1.0%
22
IEEE Access
35 papers in training set
Top 1%
1.0%
23
PeerJ
308 papers in training set
Top 10%
0.9%
24
MethodsX
16 papers in training set
Top 0.2%
0.9%
25
ACS Synthetic Biology
287 papers in training set
Top 2%
0.9%
26
BMC Methods
15 papers in training set
Top 0.1%
0.9%
27
Optics Express
26 papers in training set
Top 0.3%
0.6%
28
Nature Protocols
33 papers in training set
Top 0.5%
0.6%
29
Water Research
79 papers in training set
Top 1%
0.6%
30
Epidemics
116 papers in training set
Top 2%
0.5%