PED-X-Bench: A Benchmark of Adult-to-Pediatric Extrapolation Decisions in FDA Drug Labels
Srinivasan, A.; Berkowitz, J. S.; Friedrich, N.; Tsang, K.; Kuchi, A.; Acitores Cortina, J. M.; Zietz, M.; Czarny, R.; Liu, H.; Tatonetti, N. P.
Show abstract
Pediatric trials are ethically and logistically difficult, so the U.S. FDA often extrapolates adult data to children when justified. Yet no public resource systematically documents these decisions. We present PED-X-Bench, the first dataset and benchmark that encodes FDA pediatric-extrapolation outcomes as a four-way classification task (Full, Partial, None, Unlabeled). PED-X-Bench contains 737 FDA drug-label sections ({approx} 1 M words of source text) for approvals issued 2007-2024 across all therapeutic areas. A two-stage o3-mini prompting pipeline mined full FDA label text; nine domain reviewers then adjudicated a stratified sample of 135 labels yielding an accuracy F1 of 0.74 and 0.63 respectively (inter-annotator {kappa} = 0.678) and spot-checking the remainder. For every drug we release the ground-truth label, concise efficacy and pharmacokinetic/safety summaries, and harmonized study metadata. To showcase utility we release two baseline models: (i) a logistic-regression classifier that uses structured metadata from FDAs pediatric-labeling dataset, and (ii) a fine-tuned BigBird BERT that ingests full label text. Both base-lines perform modestly, leaving ample headroom for future work. PED-X-Bench enables research on pediatric drug development, clinical NLP and drug safety; dataset card and code are made available here: github.com/tatonetti-lab/PedXBench huggingface.co/datasets/apoorvasrinivasan/Ped-X-Bench
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A modular pipeline for natural language processing-screened human abstraction of a pragmatic trial outcome from electronic health records 93%
- Evidence Supporting EMA Drug Approvals (2020-2023): A Cross-Sectional Study of Trial Design and Outcomes 91%
- Dynamic methods for ongoing assessment of site-level risk in risk-based monitoring of clinical trials: a scoping review 90%
Similar papers in this journal
Similar papers in this journal
- Analysis of clinical trial registry entry histories using the novel R package cthist 92%
- Combining explainable machine learning, demographic and multi-omic data to identify precision medicine strategies for inflammatory bowel disease 91%
- Repurposing Therapeutics for COVID-19: Rapid Prediction of Commercially available drugs through Machine Learning and Docking 90%
Similar papers in this journal
- Controlled evaLuation of Angiotensin Receptor Blockers for COVID-19 respIraTorY disease (CLARITY): Statistical analysis plan for a randomised controlled Bayesian adaptive sample size trial 89%
- Trials that turn from retrospectively registered to prospectively registered: A cohort study of ‘retroactively prospective’ clinical trial registration using history data 89%
- Machine learning for randomised controlled trials: identifying treatment effect heterogeneity with strict control of type I error 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.