Just Speak It: Minimize Cognitive Load for Eyes-Free Text Editing with a Smart Voice Assistant
Jiayue Fan, Chenning Xu, Chun Yu, Yuanchun Shi
摘要
Entering text precisely by voice, users might encounter colloquial inserts, inappropriate wording, and recognition errors, which brings difficulties to voice editing. Users need to locate the errors and then correct them. In eyes-free scenarios, this select-modify mode brings a cognitive burden and a risk of error. This paper introduces neural networks and pre-trained models to understand users’ revision intention based on semantics, reducing the need for the information from users’ statements. We present two strategies. One is to remove the colloquial inserts automatically. The other is to allow users to edit by just speaking out the target words without having to say the context and the incorrect text. Accordingly, our approach can predict whether to insert or replace, the incorrect text to replace, and the position to insert. We implement these strategies in SmartEdit, an eyes-free voice input agent controlled with earphone buttons. The evaluation shows that our techniques reduce the cognitive load and decrease the average failure rate by 54.1% compared to descriptive command or re-speaking.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Rambler: Supporting Writing With Speech via LLM-Assisted Gist ManipulationSusan Lin, Jeremy Warner, J. D. Zamfirescu-Pereira, Matthew G. Lee 等CHI 2024 · 被引用 38 次
- GPTVoiceTasker: Advancing Multi-step Mobile Task Efficiency Through Dynamic Interface Exploration and LearningMinh Duc Vu, Han Wang, Jieshan Chen, Zhuang Li 等UIST 2024 · 被引用 17 次
- Typist Experiment: an Investigation of Human-to-Human Dictation via Role-play to Inform Voice-based Text AuthoringCan Liu, Siying Hu, Li Feng, Mingming FanCSCW 2022 · 被引用 7 次
它引用的顶会 Paper3
- Spelling Error Correction with Soft-Masked BERTShaohua Zhang, Haoran Huang, Jicong Liu, Hang LiACL 2020 · 被引用 204 次
- EYEditor: Towards On-the-Go Heads-Up Text Editing Using Voice and Manual InputDebjyoti Ghosh, Pin Sym Foong, Shengdong Zhao, Can Liu 等CHI 2020 · 被引用 42 次
- Leveraging Error Correction in Voice-based Text Entry by Talk-and-GazeKorok Sengupta, Sabin Bhattarai, Sayan Sarcar, I. Scott MacKenzie 等CHI 2020 · 被引用 14 次
相关 Paper
- Platform for Studying Self-Repairing Auto-Corrections in Mobile Text Entry based on Brain Activity, Gaze, and ContextFelix Putze, Tilman Ihrig, Tanja Schultz, Wolfgang StuerzlingerCHI 2020 · 被引用 10 次
- Voice and Touch Based Error-tolerant Multimodal Text Editing and Correction for SmartphonesMaozheng Zhao, Wenzhe Cui, I. V. Ramakrishnan, Shumin Zhai 等UIST 2021 · 被引用 17 次
- SmartFreeEdit: Mask-Free Spatial-Aware Image Editing with Complex Instruction UnderstandingQianqian Sun, Jixiang Luo, Dell Zhang, Xuelong LiACM MM 2025
- Toward Interactive DictationBelinda Z. Li, Jason Eisner, Adam Pauls, Sam ThomsonACL 2023 · 被引用 2 次
- X2T: Training an X-to-Text Typing Interface with Online Learning from User FeedbackJensen Gao, Siddharth Reddy, Glen Berseth, Nicholas Hardy 等ICLR 2021 · 被引用 10 次
