Why Invariant Risk Minimization Fails on TabularData: A Gradient Variance Solution
Mboya, G. O.
Show abstract
Machine learning models trained on observational data from one environment frequently fail when deployed in another, because standard learning algorithms exploit spurious correlations alongside causal ones. Invariant learning methods address this problem by seeking representations that support stable prediction across training environments, but their behavior on tabular data remains poorly characterized. We present CO_SCPLOWAUSC_SCPLOWTO_SCPLOWABC_SCPLOW, a gradient variance regularization framework for causal invariant representation learning on mixed tabular data. CO_SCPLOWAUSC_SCPLOWTO_SCPLOWABC_SCPLOW penalizes the variance of parameter gradients across training environments, providing a richer invariance signal than the scalar penalty used by Invariant Risk Minimization (IRM). We provide formal results showing that the gradient variance penalty is zero at causally invariant solutions and positive at solutions that rely on spurious features. Through experiments on synthetic data across three spurious-correlation regimes, four cycles of the National Health and Nutrition Examination Survey (NHANES), and four hospital systems in the UCI Heart Disease dataset, we demonstrate that: (1) IRM consistently degrades relative to standard empirical risk minimization (ERM) on tabular data, losing up to 13.8 AUC points in spurious-dominant settings, a failure we trace mechanistically to penalty collapse during training; (2) CO_SCPLOWAUSC_SCPLOWTO_SCPLOWABC_SCPLOW matches or exceeds ERM in every experimental condition; (3) CO_SCPLOWAUSC_SCPLOWTO_SCPLOWABC_SCPLOW achieves consistently better probability calibration than both ERM and IRM; and (4) invariant learning methods fail when environments differ in outcome prevalence rather than in spurious feature correlations, a boundary condition we characterize both empirically and theoretically. We introduce the Spurious Dominance Index (SDI), a practical scalar diagnostic for determining whether a dataset requires invariant learning, and validate it across all experimental settings.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Adversarial Deconfounding Autoencoder for Learning Robust Gene Expression Embeddings 95%
- Neural Collective Matrix Factorization for Integrated Analysis of Heterogeneous Biomedical Data 95%
- High-dimensional Biomarker Identification for Scalable and Interpretable Disease Prediction via Machine Learning Models 95%
Similar papers in this journal
- Compressive Big Data Analytics: An Ensemble Meta-Algorithm for High-dimensional Multisource Datasets 94%
- Voting-based integration algorithm improves causal network learning from interventional and observational data: an application to cell signaling network inference 94%
- Analyzing Biomarker Discovery: Estimating the Reproducibility of Biomarker Sets 93%
Similar papers in this journal
- Recurrent neural networks learn robust representations by dynamically balancing compression and expansion 93%
- Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports 92%
- Estimating Treatment Effects for Time-to-Treatment Antibiotic Stewardship in Sepsis 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.