N-Power AI: A Specialized Agent Framework for Automated Sample Size and Power Analysis in Clinical Trial Design
Ruan, P.; Villanueva-Miranda, I.; Liu, J.; Yang, D. M.; Zhou, Q.; Xiao, G.; Xie, Y.
Show abstract
BackgroundSample size and power analysis are essential in biomedical research and investigations, particularly in clinical trial design, as they ensure sufficient statistical power to detect meaningful effects. However, the complexity of these calculations often requires specialized statistical expertise, making the process inconvenient and limiting accessibility for researchers during early-stage study planning. MethodsWe developed N-Power AI, an agentic framework leveraging large language models (LLMs) to perform sample size and power calculations across diverse study designs. The framework consists of three specialized agents: the Function Agent, a perception and reasoning module that identifies appropriate statistical tests and corresponding R functions; the Calculation Agent, an action module that extracts parameters and executes precise computations; and the Reporting Agent, a presentation module that generates comprehensive, downloadable reports. N-Power AI and advanced LLMs (e.g., GPT-o1, Claude 3.5, Gemini 1.5 Pro) were evaluated against ground truths from statistical software (R) across six common clinical trial scenarios. ResultsDirect LLM outputs showed significant deviations from ground-truth values, particularly in complex scenarios like the Chi-Square Test and Cox Proportional Hazards Model. In contrast, N-Power AI achieved 100% agreement with ground truths across all scenarios. This accuracy is attributed to the Function Agents correct selection of statistical methods, the Calculation Agents accurate computations, and the Reporting Agents ability to produce clear and comprehensive summaries. ConclusionN-Power AI automates sample size and power analysis, offering an effective, efficient, and accessible solution for early-stage study planning. While human expertise remains crucial for high-level statistical planning, N-Power AI enhances accessibility and efficiency, streamlining the analysis process to generate reliable and reproducible results for a wide range of research scenarios.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Supporting Reanalysis and Reuse of Clinical Trial Data: A Case Study 93%
- Controlled evaLuation of Angiotensin Receptor Blockers for COVID-19 respIraTorY disease (CLARITY): Statistical analysis plan for a randomised controlled Bayesian adaptive sample size trial 93%
- Machine learning for randomised controlled trials: identifying treatment effect heterogeneity with strict control of type I error 92%
Similar papers in this journal
Similar papers in this journal
- Analysis of clinical trial registry entry histories using the novel R package cthist 95%
- An interactive retrieval system for clinical trial studies with context-dependent protocol elements 93%
- Using numerical modelling and simulation to assess the ethical burden in clinical trials and how it relates to the proportion of responders in a trial sample 92%
Similar papers in this journal
- ZIBGLMM: Zero-Inflated Bivariate Generalized Linear Mixed Model for Meta-Analysis with Double-Zero-Event Studies 94%
- Evaluation of statistical methods used to meta-analyse results from interrupted time series studies: a simulation study 94%
- Treatment recommendations based on Network Meta-Analysis: rules for risk-averse decision-makers 93%
Similar papers in this journal
- Quantitative bias analysis in practice: Review of software for regression with unmeasured confounding 93%
- Comparing randomized trial designs to estimate treatment effect in rare diseases with longitudinal models: a simulation study showcased by Autosomal Recessive Cerebellar Ataxias using the SARA score 93%
- External control arm analysis: an evaluation of propensity score approaches, G-computation, and doubly debiased machine learning 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.