The ORBIT India Dataset: Understanding the Challenges of Collecting a Disability-First AI Dataset in Low-Resource Environments
Gesu India, Martin Grayson, Cecily Morrison, Daniela Massiceti, Simon Robinson, Jennifer Pearson, Matt Jones
Abstract
Computer vision systems are increasingly used by blind individuals to navigate their lives, helping, for example, locate objects such as doors or chairs. Yet these recognition systems do not work for many personal objects a blind user might want to find, such as keys or a special notebook. In response, efforts created personalized recognition systems, where individuals train their phones to identify and locate things, like a coffee mug or white cane, using example images/videos. However, these tools are trained on data from high-resource contexts, not necessarily reflecting India’s material culture. This paper discusses the contribution of the ORBIT-India dataset, which extends these tools to the Indian context, home of the world’s largest blind population. The ORBIT-India dataset comprises 105,243 images from 587 videos, representing 76 unique objects. We use this experience to examine dataset collection practices translated from high- to low-resource settings, providing recommendations to support cross-geography dataset collection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 408906a6-78b7-496a-ade3-67ec80bd4e2bBuilds on13
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- LIMA: Less Is More for AlignmentChunting Zhou, Pengfei Liu, Puxin Xu, Srinivasan Iyer et al.NeurIPS 2023 · 1,486 citations
- Do Datasets Have Politics? Disciplinary Values in Computer Vision Dataset DevelopmentMorgan Klaus Scheuerman, Alex Hanna, Emily DentonCSCW 2021 · 169 citations
- ReCog: Supporting Blind People in Recognizing Personal ObjectsDragan Ahmetovic, Daisuke Sato, Uran Oh, Tatsuya Ishihara et al.CHI 2020 · 61 citations
- ORBIT: A Real-World Few-Shot Dataset for Teachable Object RecognitionDaniela Massiceti, Luisa M. Zintgraf, John Bronskill, Lida Theodorou et al.ICCV 2021 · 55 citations
Related papers
- Exploring the Experiences of Individuals Who are Blind or Low-Vision Using Object-Recognition Technologies in IndiaGesu India, Simon Robinson, Jennifer Pearson, Cecily Morrison et al.CHI 2025 · 8 citations
- VisAssist: A Visually Impaired-Captured Video Question Answering Benchmark for Assistive SystemsQi Gao, Heng Li, Yixin Zhou, Meixuan Zhou et al.AAAI 2026
- Disability-First Design and Creation of A Dataset Showing Private Visual Information Collected With People Who Are BlindTanusree Sharma, Abigale Stangl, Lotus Zhang, Yu-Yun Tseng et al.CHI 2023 · 23 citations
- A New Dataset Based on Images Taken by Blind People for Testing the Robustness of Image Classification Models Trained for ImageNet CategoriesReza Akbarian Bafghi, Danna GurariCVPR 2023
- Contextual Scaffolding and Self-Efficacy: Supporting Computer Skill Development among Blind Learners in IndiaAkshay Kolgar Nayak, Yash Prakash, Sampath Jayarathna, Hae Na Lee et al.CHI 2026 · 1 citation
