Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech
Katrin Tomanek, Vicky Zayats, Dirk Padfield, Kara Vaillancourt, Fadi Biadsy
Abstract
Automatic Speech Recognition (ASR) systems are often optimized to work best for speakers with canonical speech patterns. Unfortunately, these systems perform poorly when tested on atypical speech and heavily accented speech. It has previously been shown that personalization through model fine-tuning substantially improves performance. However, maintaining such large models per speaker is costly and difficult to scale. We show that by adding a relatively small number of extra parameters to the encoder layers via socalled residual adapter, we can achieve similar adaptation gains compared to model finetuning, while only updating a tiny fraction (less than 0.5%) of the model parameters. We demonstrate this on two speech adaptation tasks (atypical and accented speech) and for two state-of-the-art ASR architectures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2a376157-bf47-4f66-a038-93dffa326b13Cited by top-tier papers4
- Disentangling Voice and Content with Self-Supervision for Speaker RecognitionTianchi Liu, Kong Aik Lee, Qiongqiong Wang, Haizhou LiNeurIPS 2023 · 53 citations
- Efficient Computation Sharing for Multi-Task Visual Scene UnderstandingSara Shoouri, Mingyu Yang, Zichen Fan, Hun-Seok KimICCV 2023 · 9 citations
- Idiosyncratic Versus Normative Modeling of Atypical Speech Recognition: Dysarthric Case StudiesVishnu Raja, Adithya V. Ganesan, Anand Syamkumar, Ritwik Banerjee et al.EMNLP 2025 · 2 citations
- Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech RecognitionWonjun Lee, Hyounghun Kim, Gary LeeACL 2026
Related papers
- AdaSpeech: Adaptive Text to Speech for Custom VoiceMingjian Chen, Xu Tan, Bohan Li, Yanqing Liu et al.ICLR 2021 · 79 citations
- Accented Speech Recognition With Accent-specific CodebooksDarshan Prabhu, Preethi Jyothi, Sriram Ganapathy, Vinit UnniEMNLP 2023 · 5 citations
- A Unified Speaker Adaptation Approach for ASRYingzhu Zhao, Chongjia Ni, Cheung-Chi Leung, Shafiq R. Joty et al.EMNLP 2021
- SumRA: Parameter Efficient Fine-tuning with Singular Value Decomposition and Summed Orthogonal BasisKwok Chin Yuen, Yongsen Zheng, Jia Qi Yip, Kwok-Yan Lam et al.ICLR 2026
- Neutral residues: revisiting adapters for model extensionFranck Signe Talla, Edouard Grave, Hervé JégouICML 2025
