X2T: Training an X-to-Text Typing Interface with Online Learning from User Feedback
Jensen Gao, Siddharth Reddy, Glen Berseth, Nicholas Hardy, Nikhilesh Natraj, Karunesh Ganguly, Anca D. Dragan, Sergey Levine
摘要
We aim to help users communicate their intent to machines using flexible, adaptive interfaces that translate arbitrary user input into desired actions. In this work, we focus on assistive typing applications in which a user cannot operate a keyboard, but can instead supply other inputs, such as webcam images that capture eye gaze or neural activity measured by a brain implant. Standard methods train a model on a fixed dataset of user inputs, then deploy a static interface that does not learn from its mistakes; in part, because extracting an error signal from user behavior can be challenging. We investigate a simple idea that would enable such interfaces to improve over time, with minimal additional effort from the user: online learning from user feedback on the accuracy of the interface's actions. In the typing domain, we leverage backspaces as feedback that the interface did not perform the desired action. We propose an algorithm called x-to-text (X2T) that trains a predictive model of this feedback signal, and uses this model to fine-tune any existing, default interface for translating user input into actions that select words or characters. We evaluate X2T through a small-scale online user study with 12 participants who type sentences by gazing at their desired words, a large-scale observational study on handwriting samples from 60 users, and a pilot study with one participant using an electrocorticography-based brain-computer interface. The results show that X2T learns to outperform a non-adaptive default interface, stimulates user co-adaptation to the interface, personalizes the interface to individual users, and can leverage offline data collected from the default interface to improve its initial performance and accelerate online learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Neural Data Transformer 2: Multi-context Pretraining for Neural Spiking ActivityJoel Ye, Jennifer L. Collinger, Leila Wehbe, Robert A. GauntNeurIPS 2023 · 被引用 100 次
- First Contact: Unsupervised Human-Machine Co-Adaptation via Mutual Information MaximizationSiddharth Reddy, Sergey Levine, Anca D. DraganNeurIPS 2022 · 被引用 18 次
- Interaction-Grounded Learning with Action-Inclusive FeedbackTengyang Xie, Akanksha Saran, Dylan J. Foster, Lekan P. Molu 等NeurIPS 2022 · 被引用 12 次
- Personalized Reward Learning with Interaction-Grounded Learning (IGL)Jessica Maghakian, Paul Mineiro, Kishan Panaganti, Mark Rucker 等ICLR 2023
相关 Paper
- How We Type with Word Suggestions: Understanding Visual Attention and Checking Behavior during Mobile Text InputYang Li, Anna Maria FeitUbiComp 2025 · 被引用 3 次
- SkiMR: Dwell-free Eye Typing in Mixed RealityJinghui Hu, John J. Dudley, Per Ola KristenssonIEEE VR 2024 · 被引用 6 次
- Platform for Studying Self-Repairing Auto-Corrections in Mobile Text Entry based on Brain Activity, Gaze, and ContextFelix Putze, Tilman Ihrig, Tanja Schultz, Wolfgang StuerzlingerCHI 2020 · 被引用 10 次
- SwEYEpinch and Beyond: Exploring Intuitive, Efficient Text Entry for Extended Reality via Eye and Hand TrackingZiheng 'Leo' Li, Xichen He, Mengyuan Wu, Zeyi Tong 等CHI 2026 · 被引用 1 次
- Eye-Hand Typing: Eye Gaze Assisted Finger Typing via Bayesian Processes in ARYunlei Ren, Yan Zhang, Zhitao Liu, Ning XieIEEE VR 2024 · 被引用 10 次
