[edit]
Counterfactual Normalization: Proactively Addressing Dataset Shift Using Causal Mechanisms
Proceedings of the 34th Conference on Uncertainty in Artificial Intelligence, PMLR R16:946-956, 2018.
Abstract
Predictive models can fail to generalize from training to deployment environments because of dataset shift, posing a threat to model re- liability in practice. As opposed to previous methods which use samples from the target distribution to reactively correct dataset shift, we propose using graphical knowledge of the causal mechanisms relating variables in a pre- diction problem to proactively remove variables that participate in spurious associations with the prediction target, allowing models to gen- eralize across datasets. To accomplish this, we augment the causal graph with latent counter- factual variables that account for the underlying causal mechanisms, and show how we can es- timate these variables. In our experiments we demonstrate that models using good estimates of the latent variables instead of the observed variables transfer better from training to tar- get domains with minimal accuracy loss in the training domain.