MetaBeeAI: an AI pipeline for full-text systematic reviews in biology
Parkinson, R. H.; Cerbone, H.; Mieskolainen, M.; Cao, S.; Wilson, A. D.; Albacete, S.; Armstrong, E. B.; Bass, C.; Botias, C.; Brown, A.; Hayward, A. J.; Herbertsson, L.; Jones, A. K.; Nagloo, N.; Nicholls, E.; Rigosi, E.; Sgolastra, F.; Siviter, H.; Stanley, D. A.; Straub, L.; Straw, E. A.; Tadei, R.; Walter, K.; Stevance, H. F.; Daniels, R. K.; Lambert, B.; Roberts, S.
Show abstract
The volume and complexity of scientific literature are expanding rapidly, making it increasingly difficult to extract and synthesise information across studies. This challenge is particularly acute in the biological sciences, where evidence spans multiple levels of organisation and heterogeneous experimental designs. Large Language Model (LLM) pipelines offer a scalable route to evidence synthesis, but many existing approaches lack transparency, modularity, and effective mechanisms for human oversight. We present MetaBeeAI, an open-source, modular pipeline that integrates established LLM techniques into a coherent, auditable workflow for structured data extraction in biology. MetaBeeAI combines modular prompting, multi-pass extraction, and expert-in-the-loop validation within an interface that presents model outputs alongside source text, enabling inspection, correction, and iterative refinement. The pipeline produces machine-readable records of prompts, configurations, and expert annotations, supporting reproducibility and continuous improvement. We apply MetaBeeAI to 924 research papers on bees and pesticides, extracting structured information on species, compounds, exposure designs, and experimental context. Evaluation demonstrates improved consistency, convergence with expert judgement, and robustness across heterogeneous biological studies, highlighting the value of expert-guided refinement. MetaBeeAI provides a transparent and extensible framework for scalable evidence synthesis, supporting reliable integration of LLMs into biological research workflows.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Extraction of biological terms using large language models enhances the usability of metadata in the BioSample database 93%
- Linking big biomedical datasets to modular analysis with Portable Encapsulated Projects 92%
- Machine Learning Made Easy (MLme): A Comprehensive Toolkit for Machine Learning-Driven Data Analysis 91%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.