Demystifying Reward Design in Reinforcement Learning for Upper Extremity Interaction: Practical Guidelines for Biomechanical Simulations in HCI
Hannah Selder, Florian Fischer, Per Ola Kristensson, Arthur Fleig
摘要
Designing effective reward functions is critical for reinforcement learning-based biomechanical simulations, yet HCI researchers and practitioners often waste (computation) time with unintuitive trial-and-error tuning. This paper demystifies reward function design by systematically analyzing the impact of effort minimization, task completion bonuses, and target proximity incentives on typical HCI tasks such as pointing, tracking, and choice reaction. We show that proximity incentives are essential for guiding movement, while completion bonuses ensure task success. Effort terms, though optional, help refine motion regularity when appropriately scaled. We perform an extensive analysis of how sensitive task success and completion time depend on the weights of these three reward components. From these results we derive practical guidelines to create plausible biomechanical simulations without the need for reinforcement learning expertise, which we then validate on remote control and keyboard typing tasks. This paper advances simulation-based interaction design and evaluation in HCI by improving the efficiency and applicability of biomechanical user modeling for real-world interface development.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Learning to Utilize Shaping Rewards: A New Approach of Reward ShapingYujing Hu, Weixun Wang, Hangtian Jia, Yixiang Wang 等NeurIPS 2020 · 被引用 256 次
- The Perils of Trial-and-Error Reward Design: Misdesign through Overfitting and Invalid Task SpecificationsSerena Booth, W. Bradley Knox, Julie Shah, Scott Niekum 等AAAI 2023 · 被引用 103 次
- Predicting Mid-Air Interaction Movements and Fatigue Using Deep Reinforcement LearningNoshaba Cheema, Laura A. Frey-Law, Kourosh Naderi, Jaakko Lehtinen 等CHI 2020 · 被引用 66 次
- Explicable Reward Design for Reinforcement Learning AgentsRati Devidze, Goran Radanovic, Parameswaran Kamalaruban, Adish SinglaNeurIPS 2021 · 被引用 60 次
- Latent exploration for Reinforcement LearningAlberto Silvio Chiappa, Alessandro Marin Vargas, Ann Zixiang Huang, Alexander MathisNeurIPS 2023 · 被引用 41 次
相关 Paper
- Breathing Life Into Biomechanical User ModelsAleksi Ikkala, Florian Fischer, Markus Klar, Miroslav Bachinski 等UIST 2022 · 被引用 33 次
- Log2Motion: Biomechanical Motion Synthesis from Touch LogsMichal Patryk Miazga, Hannah Bussmann, Antti Oulasvirta, Patrick EbelCHI 2026 · 被引用 2 次
- QuestEnvSim: Environment-Aware Simulated Motion Tracking from Sparse SensorsSunmin Lee, Sebastian Starke, Yuting Ye, Jungdam Won 等SIGGRAPH 2023 · 被引用 31 次
- A Simulation Model of Intermittently Controlled Point-and-Click BehaviourSeungwon Do, Minsuk Chang, Byungjoo LeeCHI 2021 · 被引用 31 次
- Touchscreen Typing As Optimal Supervisory ControlJussi Jokinen, Aditya Acharya, Mohammad Uzair, Xinhui Jiang 等CHI 2021 · 被引用 105 次
