LLM-assisted evidence audit of late-stage cancer incidence as a screening trial endpoint
Li, S.; Zhang, W.; Xing, X.; Shen, Z.; Wang, Y.; Chen, Z.; Neto, O.; Yu, Y.; Wu, C.; Lin, L.
Show abstract
Background Late-stage cancer incidence is being considered as an earlier endpoint in cancer-screening trials, but its trial-level association with cancer-specific mortality may depend on evidence selection and endpoint harmonization. We evaluated the robustness of this association to source-verified additions. Methods We reconstructed the PubMed corpus underlying a 41-comparison review. Gemini 3.1 Pro Preview was used only to prioritize reports for blinded human reassessment. Reviewers determined eligibility, linked reports from the same trial, harmonized endpoints, and verified comparison-level data. We recalculated unweighted Pearson correlations overall and by cancer type after adding earliest-compatible trial comparisons. Results Among 1209 candidate records, 996 PDFs were assessed. Thirty-three reports absent from the source review were prioritized; 26 were eligible, representing 18 trials, and 8 provided compatible comparisons. Adding these comparisons increased the dataset from 41 to 49 and attenuated the overall correlation from 0.73 (95% confidence interval [CI] = 0.55 to 0.85) to 0.59 (95% CI = 0.37 to 0.75). Updated correlations were 0.49 (95% CI = -0.26 to 0.87) for breast, -0.23 (95% CI = -0.71 to 0.40) for colorectal, and 0.83 (95% CI = 0.54 to 0.95) for lung cancer. One sparse-event comparison influenced the colorectal estimate. Conclusions The overall association was sensitive to evidence composition, and cancer-specific stability varied. Late-stage incidence should be evaluated by cancer type and with prespecified sensitivity analyses for evidence selection and endpoint definitions. Model-assisted prioritization cannot replace human eligibility review, trial reconciliation, and source verification.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Strength of Statistical Evidence for the Efficacy of Cancer Drugs: A Bayesian Re-Analysis of Trials Supporting FDA Approval 94%
- Results dissemination from clinical trials conducted at German university medical centres was delayed and incomplete 93%
- Results reporting for clinical trials led by medical universities and university hospitals in the Nordic countries was often missing or delayed 92%
Similar papers in this journal
- Exploring scalable assessment methods for terminated trials in ClinicalTrials.gov: A cohort analysis of German and Californian trials 94%
- A Core Outcome Set to evaluate the impact of prognostication in people living with advanced cancer: an international consensus study 92%
- How Informative Were Early SARS-CoV-2 Treatment and Prevention Trials? A longitudinal cohort analysis of trials registered on clinicaltrials.gov 92%
Similar papers in this journal
- Missing data in the medical record for oncology patients: prevalence and association with outcomes 90%
- Low adherence to existing model reporting guidelines by commonly used clinical prediction models 90%
- A comprehensive systematic review and meta-analysis of the global data involving 61,532 cancer patients with SARS-CoV-2 infection 90%
Similar papers in this journal
- UKCTOCS Update: Applying insights of delayed effects in cancer screening trials to the long-term follow-up mortality analysis 94%
- Trials that turn from retrospectively registered to prospectively registered: A cohort study of ‘retroactively prospective’ clinical trial registration using history data 93%
- Supporting Reanalysis and Reuse of Clinical Trial Data: A Case Study 92%
Similar papers in this journal
- Machine learning for identifying relevant publications in updates of systematic reviews of diagnostic test studies 93%
- Exploring the potential of Claude 2 for risk of bias assessment: Using a large language model to assess randomized controlled trials with RoB 2 93%
- Evaluation of the sensitivity, accuracy and currency of the Cochrane COVID-19 Study Register for supporting rapid evidence synthesis production 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.