Gender Slopes: Counterfactual Fairness for Computer Vision Models by Attribute Manipulation
Gender Slopes: Counterfactual Fairness for Computer Vision Models by Attribute Manipulation
复制标题
DOI:
10.1145/3422841.3423533
复制
发表时间:
2020-05
期刊:
影响因子:
--
通讯作者:
Jungseock Joo;Kimmo Kärkkäinen
中科院分区:
文献类型:
--
作者:
Jungseock Joo;Kimmo Kärkkäinen
Automated computer vision systems have been applied in many domains including security, law enforcement, and personal devices, but recent reports suggest that these systems may produce biased results, discriminating against people in certain demographic groups. Diagnosing and understanding the underlying true causes of model biases, however, are challenging tasks because modern computer vision systems rely on complex black-box models whose behaviors are hard to decode. We propose to use an encoder-decoder network developed for image attribute manipulation to synthesize facial images varying in the dimensions of gender and race while keeping other signals intact. We use these synthesized images to measure counterfactual fairness of commercial computer vision classifiers by examining the degree to which these classifiers are affected by gender and racial cues controlled in the images, e.g., feminine faces may elicit higher scores for the concept of nurse and lower scores for STEM-related concepts.