Aligning Eyes between Humans and Deep Neural Network through Interactive Attention Alignment
Yuyang Gao, Tong Steven Sun, Liang Zhao, Sungsoo Ray Hong
摘要
While Deep Neural Networks (DNNs) are deriving the major innovations through their powerful automation, we are also witnessing the peril behind automation as a form of bias, such as automated racism, gender bias, and adversarial bias. As the societal impact of DNNs grows, finding an effective way to steer DNNs to align their behavior with the human mental model has become indispensable in realizing fair and accountable models. While establishing the way to adjust DNNs to "think like humans'' is in pressing need, there have been few approaches aiming to capture how "humans would think'' when DNNs introduce biased reasoning in seeing a new instance. We propose Interactive Attention Alignment (IAA), a framework that uses the methods for visualizing model attention, such as saliency maps, as an interactive medium that humans can leverage to unveil the cases of DNN's biased reasoning and directly adjust the attention. To realize more effective human-steerable DNNs than state-of-the-art, IAA introduces two novel devices. First, IAA uses Reasonability Matrix to systematically identify and adjust the cases of biased attention. Second, IAA applies GRADIA, a computational pipeline designed for effectively applying the adjusted attention to jointly maximize attention quality and prediction accuracy. We evaluated Reasonability Matrix in Study 1 and GRADIA in Study 2 in the gender classification problem. In Study 1, we found applying Reasonability Matrix in bias detection can significantly improve the perceived quality of model attention from human eyes than not applying Reasonability Matrix. In Study 2, we found using GRADIA significantly improves (1) the human-assessed perceived quality of model attention and (2) model performance in scenarios where the training samples are limited. Based on our observation in the two studies, we present implications for future design in the problem space of social computing and interactive data annotation toward achieving a human-centered steerable AI.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- RES: A Robust Framework for Guiding Visual ExplanationYuyang Gao, Tong Steven Sun, Guangji Bai, Siyi Gu 等KDD 2022 · 被引用 29 次
- Studying How to Efficiently and Effectively Guide Models with ExplanationsSukrut Rao, Moritz Böhle, Amin Parchami-Araghi, Bernt SchieleICCV 2023 · 被引用 22 次
- VisFIS: Visual Feature Importance Supervision with Right-for-the-Right-Reason ObjectivesZhuofan Ying, Peter Hase, Mohit BansalNeurIPS 2022 · 被引用 16 次
- 3DPFIX: Improving Remote Novices' 3D Printing Troubleshooting through Human-AI Collaboration DesignNahyun Kwon, Tong Steven Sun, Yuyang Gao, Liang Zhao 等CSCW 2024 · 被引用 13 次
- Designing a Direct Feedback Loop between Humans and Convolutional Neural Networks through Local ExplanationsTong Steven Sun, Yuyang Gao, Shubham Khaladkar, Sijia Liu 等CSCW 2023 · 被引用 9 次
它引用的顶会 Paper4
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang 等ICCV 2019 · 被引用 469 次
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 被引用 436 次
- Human Factors in Model Interpretability: Industry Practices, Challenges, and NeedsSungsoo Ray Hong, Jessica Hullman, Enrico BertiniCSCW 2020 · 被引用 219 次
- Silva: Interactively Assessing Machine Learning Fairness Using CausalityJing Nathan Yan, Ziwei Gu, Hubert Lin, Jeffrey M. RzeszotarskiCHI 2020 · 被引用 53 次
相关 Paper
- FAIRER: Fairness as Decision Rationale AlignmentTianlin Li, Qing Guo, Aishan Liu, Mengnan Du 等ICML 2023 · 被引用 20 次
- VisQA: X-raying Vision and Language Reasoning in TransformersTheo Jaunet, Corentin Kervadec, Romain Vuillemot, Grigory Antipov 等IEEE VIS 2021 · 被引用 30 次
- Learning to Deceive with Attention-Based ExplanationsDanish Pruthi, Mansi Gupta, Bhuwan Dhingra, Graham Neubig 等ACL 2020 · 被引用 17 次
- Learning from Observer Gaze: Zero-Shot Attention Prediction Oriented by Human-Object Interaction RecognitionYuchen Zhou, Linkai Liu, Chao GouCVPR 2024 · 被引用 13 次
- D-BIAS: A Causality-Based Human-in-the-Loop System for Tackling Algorithmic BiasBhavya Ghai, Klaus MuellerIEEE VIS 2022 · 被引用 45 次
