Towards AI-driven Sign Language Generation with Non-manual Markers
Han Zhang, Rotem Shalev-Arkushin, Vasileios Baltatzis, Connor Gillis, Gierad Laput, Raja S. Kushalnagar, Lorna C. Quandt, Leah Findlater, Abdelkareem Bedri, Colin Lea
Abstract
Sign languages are essential for the Deaf and Hard-of-Hearing (DHH) community. Sign language generation systems have the potential to support communication by translating from written languages, such as English, into signed videos. However, current systems often fail to meet user needs due to poor translation of grammatical structures, the absence of facial cues and body language, and insufficient visual and motion fidelity. We address these challenges by building on recent advances in LLMs and video generation models to translate English sentences into natural-looking AI ASL signers. The text component of our model extracts information for manual and non-manual components of ASL, which are used to synthesize skeletal pose sequences and corresponding video frames. Our findings from a user study with 30 DHH participants and thorough technical evaluations demonstrate significant progress and identify critical areas necessary to meet user needs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0cd24b0d-1f1a-49f8-8a6c-463d38ee3c9dCited by top-tier papers5
- ImageRAG: Dynamic Image Retrieval for Reference-Guided Image GenerationRotem Shalev-Arkushin, Rinon Gal, Amit Bermano, Ohad FriedICLR 2026 · 25 citations
- ImageSentinel: Protecting Visual Datasets from Unauthorized Retrieval-Augmented Image GenerationZiyuan Luo, Yangyi Zhao, Ka Chun Cheung, Simon See et al.NeurIPS 2025 · 5 citations
- Stable Signer: Hierarchical Sign Language Generative ModelSen Fang, Yalin Feng, Hongbin Zhong, Yanxin Zhang et al.ACL 2026 · 3 citations
- Reimagining Sign Language Technologies: Analyzing Translation Work of Chinese Deaf Online Content CreatorsXinru Tang, Anne Marie PiperCHI 2026 · 2 citations
- ASL Educators' Perspectives on AI for Enhancing Student Learning in American Sign Language EducationSaad Hassan, Laleh Nourian, Caluã de Lacerda Pataca, Michelle M. Olson et al.CHI 2026 · 1 citation
Builds on26
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- T2I-Adapter: Learning Adapters to Dig Out More Controllable Ability for Text-to-Image Diffusion ModelsChong Mou, Xintao Wang, Liangbin Xie, Yanze Wu et al.AAAI 2024 · 1,641 citations
Related papers
- Social App Accessibility for Deaf SignersKelly Mack, Danielle Bragg, Meredith Ringel Morris, Maarten W. Bos et al.CSCW 2020 · 52 citations
- Signs as Tokens: A Retrieval-Enhanced Multilingual Sign Language GeneratorRonglai Zuo, Rolandos Alexandros Potamias, Evangelos Ververas, Jiankang Deng et al.ICCV 2025 · 9 citations
- Customizing Generated Signs and Voices of AI Avatars: Deaf-Centric Mixed-Reality Design for Deaf-Hearing CommunicationSi Chen, Haocong Cheng, Suzy Su, Stephanie Patterson et al.CSCW 2025 · 9 citations
- Exploring the Impact of Emotional Voice Integration in Sign-to-Speech Translators for Deaf-to-Hearing CommunicationHyunchul Lim, Minghan Gao, Franklin Mingzhe Li, Nam Anh Dang et al.CSCW 2025 · 2 citations
- Perceptions and Preferences: Deaf ASL-Signing Users' Insights on Video Elements, Styles and LayoutsKhulood Alkhudaidi, Tish Burke, Rachel Boll, Shruti Mahajan et al.CHI 2025 · 1 citation
