Comparing Zealous and Restrained AI Recommendations in a Real-World Human-AI Collaboration Task
Chengyuan Xu, Kuo-Chin Lien, Tobias Höllerer
Abstract
When designing an AI-assisted decision-making system, there is often a tradeoff between precision and recall in the AI’s recommendations. We argue that careful exploitation of this tradeoff can harness the complementary strengths in the human-AI collaboration to significantly improve team performance. We investigate a real-world video anonymization task for which recall is paramount and more costly to improve. We analyze the performance of 78 professional annotators working with a) no AI assistance, b) a high-precision "restrained" AI, and c) a high-recall "zealous" AI in over 3,466 person-hours of annotation work. In comparison, the zealous AI helps human teammates achieve significantly shorter task completion time and higher recall. In a follow-up study, we remove AI assistance for everyone and find negative training effects on annotators trained with the restrained AI. These findings and our analysis point to important implications for the design of AI assistance in recall-demanding scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Interaction Configurations and Prompt Guidance in Conversational AI for Question Answering in Human-AI TeamsJaeyoon Song, Zahra Ashktorab, Qian Pan, Casey Dugan et al.CSCW 2025 · 8 citations
- Seeing Eye to Eye: Enabling Cognitive Alignment Through Shared First-Person Perspective in Human-AI Collaboration: Seeing Eye to EyeZhuyu Teng, Pei Chen, Yichen Cai, Ruoqing Lu et al.CHI 2026 · 2 citations
- AI in the Workplace: The Impact of AI on Perceived Job Decency and MeaningfulnessKuntal Ghosh, Marc Hassenzahl, Shadan SadeghianCSCW 2026
Builds on12
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 927 citations
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok et al.CHI 2021 · 713 citations
- Interpreting Interpretability: Understanding Data Scientists' Use of Interpretability Tools for Machine LearningHarmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana et al.CHI 2020 · 541 citations
- "An Ideal Human": Expectations of AI Teammates in Human-AI TeamingRui Zhang, Nathan J. McNeese, Guo Freeman, Geoff MusickCSCW 2020 · 222 citations
Related papers
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz et al.AAAI 2021 · 185 citations
- "It Became My Buddy, But I'm Not Afraid to Disagree": A Multi-Session Study of UX Evaluators Collaborating with Conversational AI AssistantsEmily Kuang, Ehsan Jahangirzadeh Soure, Luyao Shen, Nitesh Goyal et al.CHI 2026 · 3 citations
- You Complete Me: Human-AI Teams and Complementary ExpertiseQiaoning Zhang, Matthew L. Lee, Scott A. CarterCHI 2022 · 84 citations
- Does More Advice Help? The Effects of Second Opinions in AI-Assisted Decision MakingZhuoran Lu, Dakuo Wang, Ming YinCSCW 2024 · 36 citations
- Multi-Round Human–AI Collaboration with User-Specified RequirementsSima Noorani, Shayan Kiyani, Hamed Hassani, George PappasICML 2026 · 2 citations
