Discordance in pleural mesothelioma response classification and modelling of impact on clinical trials
Cowell, G. W.; Roche, J.; Noble, C.; Stobo, D. B.; Papanastasiou, A.; Kidd, A. C.; Tsim, S.; Blyth, K. G.
Show abstract
Introduction Agreement between radiologists regarding treatment response in Pleural Mesothelioma (PM) is acknowledged to be poor, but downstream effects in clinical trials have not been quantified. Methods We performed a mixed methods study, composed of a multicentre, retrospective cohort study and in silico modelling. CT images and data were retrieved from 4 UK centres regarding chemotherapy-treated patients. Expert radiologists classified response using modified Response Evaluation Criteria In Solid Tumours criteria (mRECIST) v1.1, generating discordance rate (%) and agreement. In silico modelling simulated two-arm trials of an active therapy with intended 80% power and confidence intervals for four endpoints (objective response rate (ORR), disease control rate (DCR), progression-free survival (PFS), overall survival (OS)) covering 95% of the true effect. Actual power and endpoint coverage were modelled against mRECIST misclassification rate (a single reporter equivalent of discordance rate). Consecutive simulations varied misclassification rate from 0-100% in 1% increments, each repeated 10,000 times. Results 172 cases were included. Discordance rate was 35% (60/172), kappa=0.456. In silico modelling demonstrated reduced power and endpoint precision with increasing misclassification. At 17% misclassification, corresponding to the observed 35% discordance, power dropped from 80% to 55% for ORR, 53% for DCR, 65% for PFS and 66% for OS, with endpoint coverage reduced to 88%, 89%, 92% and 92%, respectively. 50/60 (83%) discordances reflected interpretation or measurement differences intrinsic to mRECIST. Discordance was not associated with tumour volume. Conclusions Inconsistent response classification is common in PM and substantially reduces statistical power and endpoint precision in clinical trials.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Phase 1b dose expansion and translational analyses of olaparib in combination with the oral AKT inhibitor capivasertib in recurrent endometrial, triple negative breast, and ovarian, primary peritoneal, or fallopian tube cancer 91%
- Human Papilloma Virus Circulating Cell-Free DNA Kinetics in Cervical Cancer Patients Undergoing Definitive Chemoradiation 91%
- Leveraging Longitudinal Patient-Reported Outcomes Trajectories to Predict Survival in Non-Small-Cell Lung Cancer 90%
Similar papers in this journal
- An Evidenced-Based Prior for Estimating the Treatment Effect of Phase III Randomized Trials in Oncology 93%
- Clinical activity of Mitogen-Activated Protein Kinase (MAPK) inhibitors in patients with MAP2K1 (MEK1)-mutated metastatic cancers 92%
- Clinical activity of MAPK targeted therapies in patients with non-V600 BRAF mutant tumors 91%
Similar papers in this journal
- Revisiting a null hypothesis: exploring the parameters of oligometastasis treatment 91%
- Contralateral Neck Recurrence Rates After Ipsilateral Neck Adjuvant Radiation in Head and Neck Carcinomas with a Pathologically Negative Contralateral Neck 90%
- Noninvasive Quantification of Radiation-Induced Lung Injury using a Targeted Molecular Imaging Probe 90%
Similar papers in this journal
- Reproducibility and transparency characteristics of oncology research evidence 92%
- Surgical Resection, Radiotherapy, And Percutaneous Thermal Ablation for Treatment of Stage 1 Non-Small Cell Lung Cancer: A Systematic Review and Network Meta-Analysis 91%
- Selumetinib in combination with dexamethasone for the treatment of relapsed/refractory RAS-pathway mutated paediatric and adult acute lymphoblastic leukaemia (SeluDex): study protocol for an international, parallel-group, dose-finding with expansion phase I/II trial 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.