Tap&Say: Touch Location-Informed Large Language Model for Multimodal Text Correction on Smartphones
Maozheng Zhao, Michael Xuelin Huang, Nathan G. Huang, Shanqing Cai, Henry Huang, Michael G. Huang, Shumin Zhai, I. V. Ramakrishnan, Xiaojun Bi
2025年份
6被引次数
1顶会引用
摘要
layer that integrates the tap location into the LLM's attention mechanism, enabling it to utilize the tap location for text correction. We fine-tuned the touch location-informed LLM on synthetic touch locations and correction commands, achieving significantly higher correction accuracy than the state-of-the-art method VT [45]. A 16-person user study demonstrated that Tap&Say outperforms VT [45] with 16.4% shorter task completion time and 47.5% fewer keyboard clicks and is preferred by users.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- LLM Powered Text Entry Decoding and Flexible Typing on SmartphonesYan Ma, Dan Zhang, I. V. Ramakrishnan, Xiaojun BiCHI 2025 · 被引用 1 次
- Voice and Touch Based Error-tolerant Multimodal Text Editing and Correction for SmartphonesMaozheng Zhao, Wenzhe Cui, I. V. Ramakrishnan, Shumin Zhai 等UIST 2021 · 被引用 17 次
- Desirable Unfamiliarity: Insights from Eye Movements on Engagement and Readability of Dictation InterfacesZhaohui Liang, Yonglin Chen, Naser Al Madi, Can LiuCHI 2026 · 被引用 1 次
- Leveraging Error Correction in Voice-based Text Entry by Talk-and-GazeKorok Sengupta, Sabin Bhattarai, Sayan Sarcar, I. Scott MacKenzie 等CHI 2020 · 被引用 14 次
- Text Input for Non-Stationary XR Workspaces: Investigating Tap and Word-Gesture Keyboards in Virtual and Augmented RealityFlorian Kern, Florian Niebling, Marc Erich LatoschikIEEE VR 2023 · 被引用 40 次
