Aug 2026· Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.2· pp. 5698-5707· 0 citations· 43 references
TL;DR
Concept-Residual eXpansion (CRX), a concept-augmented framework that improves robustness by expanding the set of candidate predictive features by improving robustness to spurious correlations, is proposed.
Abstract
Models trained with empirical risk minimization (ERM) are prone to relying on spurious correlations to make predictions. A spurious correlation is a non-causal relationship in the training data between an attribute and the prediction target that does not generalize beyond the training environment. As a result, models can appear to achieve strong performance by exploiting these correlations, yet fail when the correlation changes or disappears. Despite their tendency to learn spurious correlations, the success of post-hoc mitigation methods in recent work suggests that ERM-trained models still retain useful, robust predictive features. However, core (non-spurious) features may be weak or entangled within the representation, making them difficult to identify. We propose Concept-Residual eXpansion (CRX), a concept-augmented framework that improves robustness by expanding the set of candidate predictive features. Starting from a frozen ERM representation, we augment the model's features with interpretable concept scores that describe the presence of task-relevant attributes and the surrounding context, together with residual features that capture the portion of the ERM features not expressed by the concepts. We then retrain a lightweight classifier on this expanded feature space, enabling it to leverage both structured semantic cues and complementary residual information. Across standard spurious correlation benchmarks, CRX consistently improves worst-group accuracy while maintaining competitive average performance. These results suggest that expanding the set of available features can substantially improve robustness to spurious correlations. The code and additional implementation details can be found at https://doi.org/10.5281/zenodo.20467988
Deep neural networks often learn and rely on spurious correlations, i.e., superficial associations between non-causal features and the targets. For instance, an image classifier may identify camels based on the desert backgrounds. While it can yield high overall accuracy during training, it degrades generalization on m...
Wenqian Ye, Guangtao Zheng, Ai-Dong Zhang· Knowledge Discovery and Data...· 5 citations
It is shown that a usable signal is available after convergence, when loss no longer distinguishes the two populations, and applying a fixed perturbation to a converged model's inputs flips the predictions of the latter far more often than the former.
Deep neural networks tend to rely on simple features that may be spurious and thus fail to generalize. We study this problem in the setting of linear probes, where a (generalized) linear model is fitted on the representations of a (pretrained) model. We use the connection of these models to the max-margin classifier, a...
Floris Holstege, Bram Wouters, N. V. van Giersbergen et al.· 0 citations
Subpopulation-Aware Generative Enhancement (SAGE), a two-stage generative augmentation framework, is introduced, using cluster-derived sub-labels and class labels to fine-tune a conditional generative model and text encoder, generating targeted synthetic data to fill underrepresented regions in the training set and con...
Yi-Ming Luo, Rong-Qiang Zhao, Jie Liu· 0 citations
Latent world models enable planning by predicting the effects of actions in a learned representation space, but their predictions can become unreliable when test-time conditions differ from training. Existing test-time adaptation methods address this by updating parts of the pretrained model, often modifying millions o...
Krishnam Soni, Aditya Sehgal, Vedant Dave et al.· 0 citations
Deterministic neural networks and neural operators provide point predictions with no intrinsic measure of reliability. Yet, predictive uncertainty may stem from irreducible outcome variability, finite data, or limitations of the chosen model class. Monte Carlo dropout offers a computationally convenient way to construc...
Giacomo Lorenzon, Francesco Regazzoni· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.