MuSTAF: Clinically Relevant Multi-task Spatiotemporal Attention Fusion Framework for Breast Cancer Detection with Longitudinal Mammography
Li, Y.; Castelo, A.; Dennison, J. B.; Kettner, N. M.; Sieh, W.; Joseph, J. R.; Castillo, E.; Brock, K.; Weaver, O. O.; Wu, C.
Show abstract
Recent NCCN guideline highlighted AI-based mammographic risk prediction, but AI-based breast cancer detection remains questionable to translation. One barrier is current models often do not match routine clinical reasoning, which may add decision burden than benefits. In practice, radiologists compare current and prior mammograms while assessing breast density, bilateral symmetry, and lesion laterality. To align AI with this reasoning, we developed MuSTAF, a multi-task spatiotemporal attention fusion model for patient-level breast cancer classification from longitudinal full-field digital mammography. MuSTAF uses up to three recent mammograms, integrates temporal and cross-view information, refines suspicious-region features, and jointly predicts cancer status, breast density, and bilateral symmetry, with a separate laterality classifier for cancer-positive cases. In an internal case-control cohort (n = 351), MuSTAF achieved a cancer classification (AUC=0.84) exceeding all architecture-level baselines and published mammography AI models adapted to the same task (AUC [≤] 0.81). Simultaneously, it achieved AUCs of 0.83/0.80 for density/laterality assessments, and removing these auxiliary tasks reduced cancer detection performance. On the external CSAW-CC dataset (n = 8,723), model performance improved from 0.72 to 0.88 when restricting cancer cases to those with latest exams within 60 days before diagnosis, showing that temporally distant labels may shift detection evaluation toward risk prediction. Longitudinal analysis further showed that three recent exams outperformed five exams internally (AUC = 0.84 vs 0.80) and externally (0.72 vs 0.66), indicating recent imaging evidence mattered more than remote history. Overall, MuSTAF model improved longitudinal mammographic cancer classification while providing auxiliary outputs, and clarified temporal factors for applying AI to screening detection.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Dual Adversarial Deconfounding Autoencoder for joint batch-effects removal from multi-center and multi-scanner radiomics data 94%
- Reproducible And Clinically Translatable Deep Neural Networks For Cervical Screening 94%
- Spatial Transcriptomics Inferred from Pathology Whole-Slide Images Links Tumor Heterogeneity to Survival in Breast and Lung Cancer 94%
Similar papers in this journal
- Artificial Intelligence System Reduces False-Positive Findings in the Interpretation of Breast Ultrasound Exams 95%
- Generative AI Enables Medical Image Segmentation in Ultra Low-Data Regimes 93%
- Features fusion or not: harnessing multiple pathological foundation models using Meta-Encoder for downstream tasks fine-tuning 93%
Similar papers in this journal
- A human-in-the-loop explanation framework for morphologically transparent AI predictions from whole-slide images 93%
- STPath: A Generative Foundation Model for Integrating Spatial Transcriptomics and Whole Slide Images 92%
- Understanding the robustness of vision-language models to medical image artefacts 91%
Similar papers in this journal
- A Deep Learning Model for Molecular Label Transfer that Enables Cancer Cell Identification from Histopathology Images 92%
- Deep learning inference of cell type-specific gene expression from breast tumor histopathology 91%
- Image-Based Consensus Molecular Subtyping in Rectal Cancer Biopsies and Response to Neoadjuvant Chemoradiotherapy 91%
Similar papers in this journal
- Unmasking the tissue microecology of ductal carcinoma in situ with deep learning 92%
- Normal Breast Tissue (NBT)-Classifiers: Advancing Compartment Classification in Normal Breast Histology 91%
- Predicting neoadjuvant chemotherapy benefit using deep learning from stromal histology in breast cancer 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.