MAVeRiC-AD: Mixture-of-experts Agentic Vision-Language Ensemble for Robust MRI Classification of Alzheimer's Disease
Dhinagar, N. J.; Senthilkumar, P.; Thomopoulos, S. I.; Thompson, P. M.
Show abstract
Robust classification of Alzheimers disease (AD) from structural T1-weighted MRI (T1w) images remains an unmet clinical need, especially when data is acquired at multiple sites that differ in scanning protocols and population demographics. In this paper, we present MAVeRiC-AD (Mixture-of-experts Agent-guided Vision-Language Ensemble for Robust Imaging-based classification of Alzheimers Disease), an agentic framework that dynamically utilizes the optimal inferencing tool for radiological queries to provide relevant answers to the user. In our framework, we tested three specialized models encoded as callable tools: (1) CNN-AD, a 3D DenseNet trained on T1w intensities only; (2) MOE-VLM, a vision-language model that jointly models the T1w with subject-specific demographics (age, sex, site) via a mixture-of-experts (MoE) projection head; (3) Retrieval engine, a similarity-search module that contextualizes a patient against others from the site and reports % prevalence of AD. A light-weight agent analyzes the user request and then routes the input (image, or image + text) to the appropriate tool, aggregates responses and returns the tool response augmented with its confidence derived from conformal prediction. Experiments were conducted using T1w images from the ADNI (N=4,098) and OASIS-3 (N=600) datasets. Single-site training baselines achieved ROC-AUC = 0.79 (CNN) and 0.82 (VLM) on ADNI. When trained jointly on both sites, MOE-VLM surpassed both image-only and standard vision-language models with ROC-AUC = 0.90 on ADNI and 0.81 on OASIS. MAVeRiC-AD demonstrates that agentic orchestration of complementary expert deep models, coupled with explicit demographic conditioning for multi-site data can improve robustness and interpretability of AD image analysis pipelines and serves as a blueprint for scalable, trustworthy clinical AI assistants.
Matching journals
The top 13 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- AI-driven fusion of neurological work-up for assessment of biological Alzheimer’s disease 90%
- Generative AI Enables Medical Image Segmentation in Ultra Low-Data Regimes 90%
- Segmenting functional tissue units across human organs using community-driven development of generalizable machine learning algorithms 90%
Similar papers in this journal
- Cross-dataset Evaluation of Dementia Longitudinal Progression Prediction Models 94%
- WMH-DualTasker: A weakly-supervised deep learning model for automated white matter hyperintensities segmentation and visual rating prediction 94%
- DeepComBat: A Statistically Motivated, Hyperparameter-Robust, Deep Learning Approach to Harmonization of Neuroimaging Data 93%
Similar papers in this journal
- NeuroConText: Contrastive Learning for Neuroscience Meta-Analysis with Rich Text Representation 93%
- ReMiND: Recovery of Missing Neuroimaging using Diffusion Models with Application to Alzheimer’s Disease 93%
- BrainQCNet: a Deep Learning attention-based model for the automated detection of artifacts in brain structural MRI scans. 91%
Similar papers in this journal
- A Clinical Neuroimaging Platform for Rapid, Automated Lesion Detection and Personalized Post-Stroke Outcome Prediction 91%
- STPath: A Generative Foundation Model for Integrating Spatial Transcriptomics and Whole Slide Images 91%
- Interpretable deep learning approach for extracting cognitive features from hand-drawn images of intersecting pentagons in older adults 90%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.