Opportunities and Challenges of Automatic Speech Recognition Systems for Low-Resource Language Speakers
Thomas Reitmaier, Electra Wallington, Dani Kalarikalayil Raju, Ondrej Klejch, Jennifer Pearson, Matt Jones, Peter Bell, Simon Robinson
Abstract
Automatic Speech Recognition (ASR) researchers are turning their attention towards supporting low-resource languages, such as isiXhosa or Marathi, with only limited training resources. We report and reflect on collaborative research across ASR & HCI to situate ASR-enabled technologies to suit the needs and functions of two communities of low-resource language speakers, on the outskirts of Cape Town, South Africa and in Mumbai, India. We build on longstanding community partnerships and draw on linguistics, media studies and HCI scholarship to guide our research. We demonstrate diverse design methods to: remotely engage participants; collect speech data to test ASR models; and ultimately field-test models with users. Reflecting on the research, we identify opportunities, challenges, and use-cases of ASR, in particular to support pervasive use of WhatsApp voice messaging. Finally, we uncover implications for collaborations across ASR & HCI that advance important discussions at CHI surrounding data, ethics, and AI.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 94fb8477-1a15-401d-94fb-cd82a41e771bCited by top-tier papers8
- Toward Language Justice: Exploring Multilingual Captioning for AccessibilityAashaka Desai, Rahaf Alharbi, Stacy Hsueh, Richard E. Ladner et al.CHI 2025 · 11 citations
- Situating Automatic Speech Recognition Development within Communities of Under-heard Language SpeakersThomas Reitmaier, Electra Wallington, Ondrej Klejch, Nina Markl et al.CHI 2023 · 9 citations
- Low-Resourced Languages and Online Knowledge Repositories: A Need-Finding StudyHellina Hailu Nigatu, John F. Canny, Sarah E. ChasinsCHI 2024 · 8 citations
- Cultivating Spoken Language Technologies for Unwritten LanguagesThomas Reitmaier, Dani Kalarikalayil Raju, Ondrej Klejch, Electra Wallington et al.CHI 2024 · 5 citations
- The Esethu Framework: Reimagining Sustainable Dataset Governance and Curation for Low-Resource LanguagesJenalea Rajab, Anuoluwapo Aremu, Everlyn Asiko Chimoto, Dale Dunbar et al.ACL 2025 · 3 citations
Related papers
- Remotely Co-Designing Features for Communication Applications using Automatic Captioning with Deaf and Hearing PairsMatthew Seita, Sooyeon Lee, Sarah Andrew, Kristen Shinohara et al.CHI 2022 · 51 citations
- Towards Building ASR Systems for the Next Billion UsersTahir Javed, Sumanth Doddapaneni, Abhigyan Raman, Kaushal Santosh Bhogale et al.AAAI 2022 · 86 citations
- Collectively Reimagining Artificial Intelligence With Marginalized CommunitiesRachel Oluwatuyi, Vimlan Pillay, Jasmine Mazwi, Alvin Castro et al.CHI 2026 · 1 citation
- Learning From Failure: Data Capture in an Australian Aboriginal CommunityÉric Le Ferrand, Steven Bird, Laurent BesacierACL 2022
- Ethical Considerations for Machine Translation of Indigenous Languages: Giving a Voice to the SpeakersManuel Mager, Elisabeth Mager, Katharina Kann, Ngoc Thang VuACL 2023 · 15 citations
