CLARIS: Clear and Intelligible Speech from Whispered and Dysarthric Voices
Neil Shah, Yash Sonkar, Shirish Subhash Karande, Vineet Gandhi
Abstract
Whispered and dysarthric speech hinder effective communication and undermine the reliability of voice-enabled systems. We present CLARIS, a compact speech-to-speech restoration system that turns such atypical input into clear, expressive speech. CLARIS requires no disorder-specific architectural tuning, generalizes across languages, and adapts quickly to new accents and speakers, enabling practical personalization. On whispered English, Hindi, and clinically challenging dysarthric speech, CLARIS delivers state-of-the-art intelligibility and naturalness, with listener studies confirming gains in quality, intelligibility, naturalness, and prosody. The system runs in real time, converting one second of input in about 30ms and enables inclusive, private, and personalized voice interaction. Audio samples are available at https://claris-w2s.github.io/CLARIS/
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 5b3c5777-4d09-4a1e-a8a9-82c58a90227dRelated papers
- WESPER: Zero-shot and Realtime Whisper to Normal Voice Conversion for Whisper-based Speech InteractionsJun RekimotoCHI 2023 · 24 citations
- Idiosyncratic Versus Normative Modeling of Atypical Speech Recognition: Dysarthric Case StudiesVishnu Raja, Adithya V. Ganesan, Anand Syamkumar, Ritwik Banerjee et al.EMNLP 2025 · 2 citations
- EA-VAE: Learning to Reconstruct Dysarthric Speech via Variational Autoencoder with Encoding AlignmentDaipeng Zhang, Wenhuan Lu, Xianghu Yue, Hongcheng Zhang et al.AAAI 2026
- DualVoice: Speech Interaction that Discriminates between Normal and Whispered Voice InputJun RekimotoUIST 2022 · 9 citations
- From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech RecognitionTianduo Wang, Lu Xu, Wei Lu, Shanbo ChengEMNLP 2025 · 1 citation
