Explanatory Debiasing: Involving Domain Experts in the Data Generation Process to Mitigate Representation Bias in AI Systems
Aditya Bhattacharya, Simone Stumpf, Robin De Croon, Katrien Verbert
Abstract
Representation bias is one of the most common types of biases in artificial intelligence (AI) systems, causing AI models to perform poorly on underrepresented data segments. Although AI practitioners use various methods to reduce representation bias, their effectiveness is often constrained by insufficient domain knowledge in the debiasing process. To address this gap, this paper introduces a set of generic design guidelines for effectively involving domain experts in representation debiasing. We instantiated our proposed guidelines in a healthcare-focused application and evaluated them through a comprehensive mixed-methods user study with 35 healthcare experts. Our findings show that involving domain experts can reduce representation bias without compromising model accuracy. Based on our findings, we also offer recommendations for developers to build robust debiasing systems guided by our generic design guidelines, ensuring more effective inclusion of domain experts in the debiasing process.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3e7551a6-7f4b-4308-9a2f-fa378f3e3ae3Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Re-examining Whether, Why, and How Human-AI Interaction Is Uniquely Difficult to DesignQian Yang, Aaron Steinfeld, Carolyn P. Rosé, John ZimmermanCHI 2020 · 604 citations
- The Effects of Regularization and Data Augmentation are Class DependentRandall Balestriero, Léon Bottou, Yann LeCunNeurIPS 2022 · 124 citations
- Data-Centric Explanations: Explaining Training Data of Machine Learning Systems to Promote TransparencyAriful Islam Anik, Andrea BuntCHI 2021 · 99 citations
- You Complete Me: Human-AI Teams and Complementary ExpertiseQiaoning Zhang, Matthew L. Lee, Scott A. CarterCHI 2022 · 84 citations
- Data preprocessing to mitigate bias: A maximum entropy based approachL. Elisa Celis, Vijay Keswani, Nisheeth K. VishnoiICML 2020 · 45 citations
Related papers
- Improving Subgroup Robustness via Data SelectionSaachi Jain, Kimia Hamidieh, Kristian Georgiev, Andrew Ilyas et al.NeurIPS 2024 · 17 citations
- Domain Experience and Expertise in Explainable AI Applications: A Bearing Fault Diagnosis Case StudyZibin Zhao, Michael Castelle, Cagatay TurkayCSCW 2025 · 1 citation
- D-BIAS: A Causality-Based Human-in-the-Loop System for Tackling Algorithmic BiasBhavya Ghai, Klaus MuellerIEEE VIS 2022 · 45 citations
- Mitigating Sentiment Bias for Recommender SystemsChen Lin, Xinyi Liu, Guipeng Xv, Hui LiSIGIR 2021 · 31 citations
- Learning De-biased Representations with Biased RepresentationsHyojin Bahng, Sanghyuk Chun, Sangdoo Yun, Jaegul Choo et al.ICML 2020 · 332 citations
