Back

Immortal time bias reproduces the reported survival benefit of conversion surgery in stage IV gastric cancer: a simulation study

Sah, B. K.; Li, C.; Li, J.; Zhu, Z.

2026-09-03 gastroenterology
10.64898/2026.09.01.26361986 medRxiv
Show abstract

Background Conversion surgery for stage IV gastric cancer is supported by a pooled overall survival hazard ratio of 0.36 (95% confidence interval 0.32-0.40) and, in the largest international cohort, median survival of 36.7 versus 12.5-13.8 months on chemotherapy. Survival is measured from diagnosis; the median diagnosis-to-gastrectomy interval is 124 days, which patients must survive to be counted surgical. Methods We simulated cohorts of 3,177 stage IV gastric cancer patients from published parameters: background median survival 14.5 months; median diagnosis-to-surgery interval 124 days (category-specific 92-174 days). Surgery had no effect (true hazard ratio 1.00 by construction). Data were analysed as the literature analyses them (exposure fixed at baseline, follow-up from diagnosis), and by time-varying Cox and landmark analysis. Confounding by indication was added in a second scenario. Results Under immortal time bias alone the naive analysis returned a hazard ratio of 0.794 (95% simulation interval 0.743-0.851), median survival 16.8 versus 12.8 months. Time-varying Cox recovered 1.000 and landmark analysis 1.000-1.004. Bias scaled with the interval: 0.849 at 92 days, 0.715 at 174 days. Adding confounding, the naive estimate fell to 0.601 (0.560-0.644) at strength 0.5 and 0.356 (0.323-0.385) at strength 1.5, overlapping the published estimate; median survival 21.9 versus 8.7 months. Correcting immortal time alone left residual bias (hazard ratio 0.439). Conclusions The reported survival advantage of conversion surgery is reproducible where the operation does nothing; published estimates cannot distinguish benefit from bias. Resolving this requires individual patient data analysed with methods that assign person-time correctly, or completion of JCOG2301.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.