Practical Bias Mitigation through Proxy Sensitive Attribute Label Generation


この目的に向けて、教師なし埋め込み生成とそれに続くプロキシ依存ラベルを取得するためのクラスタリングという 2 段階のアプローチを提案します。
実験結果は、Fair Mixup や Adversarial Debiasing などの既存のアルゴリズムを使用したバイアス軽減が、真の機密属性を使用した場合と比較した場合、派生プロキシ ラベルで同等の結果をもたらすことを示しています。


Addressing bias in the trained machine learning system often requires access to sensitive attributes. In practice, these attributes are not available either due to legal and policy regulations or data unavailability for a given demographic. Existing bias mitigation algorithms are limited in their applicability to real-world scenarios as they require access to sensitive attributes to achieve fairness. In this research work, we aim to address this bottleneck through our proposed unsupervised proxy-sensitive attribute label generation technique. Towards this end, we propose a two-stage approach of unsupervised embedding generation followed by clustering to obtain proxy-sensitive labels. The efficacy of our work relies on the assumption that bias propagates through non-sensitive attributes that are correlated to the sensitive attributes and, when mapped to the high dimensional latent space, produces clusters of different demographic groups that exist in the data. Experimental results demonstrate that bias mitigation using existing algorithms such as Fair Mixup and Adversarial Debiasing yields comparable results on derived proxy labels when compared against using true sensitive attributes.


著者 Bhushan Chaudhary,Anubha Pandey,Deepak Bhatt,Darshika Tiwari
発行日 2023-12-26 10:54:15+00:00
arxivサイト arxiv_id(pdf)

提供元, 利用サービス, Google

カテゴリー: cs.CY, cs.LG パーマリンク