Assessing Image Quality Issues for Real-World Problems
Tai-Yin Chiu, Yinan Zhao, Danna Gurari
摘要
We introduce a new large-scale dataset that links the assessment of image quality issues to two practical vision tasks: image captioning and visual question answering. First, we identify for 39,181 images taken by people who are blind whether each is sufficient quality to recognize the content as well as what quality flaws are observed from six options. These labels serve as a critical foundation for us to make the following contributions: (1) a new problem and algorithms for deciding whether an image is insufficient quality to recognize the content and so not captionable, (2) a new problem and algorithms for deciding which of six quality flaws an image contains, (3) a new problem and algorithms for deciding whether a visual question is unanswerable due to unrecognizable content versus the content of interest being missing from the field of view, and (4) a novel application of more efficiently creating a large-scale image captioning dataset by automatically deciding whether an image is insufficient quality and so should not be captioned. We publicly-share our datasets and code to facilitate future extensions of this work: https://vizwiz.org .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Grounding Answers for Visual Questions Asked by Visually Impaired PeopleChongyan Chen, Samreen Anjum, Danna GurariCVPR 2022 · 被引用 48 次
- MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction TuningZhiyang Xu, Ying Shen, Lifu HuangACL 2023 · 被引用 37 次
- Disability-First Design and Creation of A Dataset Showing Private Visual Information Collected With People Who Are BlindTanusree Sharma, Abigale Stangl, Lotus Zhang, Yu-Yun Tseng 等CHI 2023 · 被引用 23 次
- Right this way: Can VLMs Guide Us to See More to Answer Questions?Li Liu, Diji Yang, Sijia Zhong, Kalyana Suma Sree Tholeti 等NeurIPS 2024 · 被引用 20 次
- Vision Skills Needed to Answer Visual QuestionsXiaoyu Zeng, Yanan Wang, Tai-Yin Chiu, Nilavra Bhattacharya 等CSCW 2020 · 被引用 17 次
它引用的顶会 Paper2
相关 Paper
- A New Dataset Based on Images Taken by Blind People for Testing the Robustness of Image Classification Models Trained for ImageNet CategoriesReza Akbarian Bafghi, Danna GurariCVPR 2023
- "It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language ModelsKapil Garg, Xinru Tang, Jimin Heo, Dwayne R. Morgan 等CHI 2026 · 被引用 2 次
- VisAssist: A Visually Impaired-Captured Video Question Answering Benchmark for Assistive SystemsQi Gao, Heng Li, Yixin Zhou, Meixuan Zhou 等AAAI 2026
- "I Hope This Is Helpful": Understanding Crowdworkers' Challenges and Motivations for an Image Description TaskRachel N. Simons, Danna Gurari, Kenneth R. FleischmannCSCW 2020 · 被引用 27 次
- Accessibility for Color Vision Deficiencies: Challenges and Findings of a Large Scale Study on Paper FiguresKatrin Angerbauer, Nils Rodrigues, René Cutura, Seyda Öney 等CHI 2022 · 被引用 32 次
