Current browse context:
cs.LG
Change to browse by:
References & Citations
Computer Science > Machine Learning
Title: On the Fairness of Disentangled Representations
(Submitted on 31 May 2019 (v1), last revised 29 Oct 2019 (this version, v2))
Abstract: Recently there has been a significant interest in learning disentangled representations, as they promise increased interpretability, generalization to unseen scenarios and faster learning on downstream tasks. In this paper, we investigate the usefulness of different notions of disentanglement for improving the fairness of downstream prediction tasks based on representations. We consider the setting where the goal is to predict a target variable based on the learned representation of high-dimensional observations (such as images) that depend on both the target variable and an \emph{unobserved} sensitive variable. We show that in this setting both the optimal and empirical predictions can be unfair, even if the target variable and the sensitive variable are independent. Analyzing the representations of more than \num{12600} trained state-of-the-art disentangled models, we observe that several disentanglement scores are consistently correlated with increased fairness, suggesting that disentanglement may be a useful property to encourage fairness when sensitive variables are not observed.
Submission history
From: Francesco Locatello [view email][v1] Fri, 31 May 2019 15:03:12 GMT (2693kb,D)
[v2] Tue, 29 Oct 2019 10:56:08 GMT (2712kb,D)
Link back to: arXiv, form interface, contact.