Counterfactual Marginalisation for Robustness Testing
September 11, 2026
This framework evaluates model robustness by using counterfactual image generators to intervene on nuisance variables like age or sex during test-time. By averaging predictions over these interventions, the method produces stability and calibration metrics that account for demographic shortcuts.
HOW THIS AFFECTS YOU
●
researcherYou can utilize this to quantify how much your model relies on non-causal demographic features.
●
policyThis provides a mathematical framework for auditing model fairness and bias.