A software package for simple and rigorous survival machine learning analysis in biomedical research
Pybus, A.; Qiu, J.; Morais Lyra, P. C.; Dang, K.; Narvaez-Bandera, I.; Jolaogun, T.; Goecks, J.
Show abstract
Survival analysis is a fundamental technique in biomedical research for modeling time-to-event data. It enables the identification of prognostic factors in disease, compares survival outcomes across treatment groups, and performs targeted treatment selection. A variety of machine learning (ML) approaches to survival analysis have emerged to complement classical statistical methods, especially for high-dimensional datasets with complex, nonlinear interactions between features. However, using survival ML methods requires addressing challenges such as censoring-unaware evaluation, overfitting, selecting performance metrics, and data leakage. To address these and other difficulties in using survival ML models, we developed the mlsurv software package. mlsurv is an open-source Python package built around three major design principles: 1) methodological rigor, including evidence-based model selection, leakage-free pipelines, and multi-metric evaluation, 2) multi-scale evaluation and interpretation, including population and subpopulation evaluation, patient-level explanations, and feature analysis, and 3) automated trust and transparency, including limitation flagging and TRIPOD+AI-aligned reporting. mlsurv bundles ten models spanning linear, ensemble, kernel, and deep learning families within a unified software package. We demonstrate mlsurv on the Chowell immunotherapy cohort (n=1,479). The survival-trained models achieve a test concordance index of 0.73 for overall survival prediction. Further, risk scores strongly correlate with the response-trained LORIS clinical score (|{rho}| up to 0.84), reflecting the overlap between prognostic and predictive signal. mlsurv enables biomedical researchers to conduct rigorous, multi-model survival analysis and benchmarking using minimal code with default best practices rather than implementing custom scripts and methodological safeguards from scratch.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- PiDeel: Pathway-informed deep learning model for survivalanalysis and pathological classification of gliomas 89%
- Comprehensive evaluation of cell-type quantification methods for immuno-oncology 89%
- Using Cancer Profiles to Identify Synthetic Lethal Therapeutic Targets and Predictive Biomarkers in Cancer Gene Dependency Data 88%
Similar papers in this journal
- RiskPath : Explainable deep learning for multistep biomedical prediction in longitudinal data 89%
- cytoGPNet: Enhancing Clinical Outcome Prediction Accuracy Using Longitudinal Cytometry Data in Small Cohort Studies 88%
- Federated Learning for multi-omics: a performance evaluation in Parkinson's disease 88%
Similar papers in this journal
- Cancer patient survival can be accurately parameterized, revealing time-dependent therapeutic effects and doubling the precision of small trials 91%
- Integration of clinical, pathological, radiological, and transcriptomic data improves the prediction of first-line immunotherapy outcome in metastatic non-small cell lung cancer 90%
- Uncovering interpretable potential confounders in electronic medical records 90%
Similar papers in this journal
- KMSubtraction: Reconstruction of unreported subgroup survival data utilizing published Kaplan-Meier survival curves 88%
- A Comparative Analysis in a Clinical Cohort: Multiple Imputation by Chained Equations and a Novel Super Learner-Based Imputation Approach 88%
- Prediction-powered Inference for Clinical Trials 87%
Similar papers in this journal
- Robust Evaluation of Deep Learning-based Representation Methods for Survival and Gene Essentiality Prediction on Bulk RNA-seq Data 90%
- MultiSurv: Long-term cancer survival prediction using multimodal deep learning 90%
- Leveraging Temporal Learning with Dynamic Range (TLDR) for Enhanced Prediction of Outcomes in Recurrent Exposure and Treatment Settings in Electronic Health Records 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.