Demystifying Reward Design in Reinforcement Learning for Upper Extremity Interaction: Practical Guidelines for Biomechanical Simulations in HCI
Authors
Designing effective reward functions is critical for reinforcement learning-based biomechanical simulations, yet HCI researchers and practitioners often waste (computation) time with unintuitive trial-and-error tuning. This paper demystifies reward function design by systematically analyzing the impact of effort minimization, task completion bonuses, and target proximity incentives on typical HCI tasks such as pointing, tracking, and choice reaction. We show that proximity incentives are essential for guiding movement, while completion bonuses ensure task success. Effort terms, though optional, help refine motion regularity when appropriately scaled. We perform an extensive analysis of how sensitive task success and completion time depend on the weights of these three reward components. From these results we derive practical guidelines to create plausible biomechanical simulations without the need for reinforcement learning expertise, which we then validate on remote control and keyboard typing tasks. This paper advances simulation-based interaction design and evaluation in HCI by improving the efficiency and applicability of biomechanical user modeling for real-world interface development.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 83%
Log2Motion: Biomechanical Motion Synthesis from Touch Logs
CHI '26· Hand Gesture Recognition +2
- 80%
Weak-Annotation of HAR Datasets using Vision Foundation Models
UbiComp '24· Human Pose & Activity Recognition +1
- 80%
Servo-Gaussian Model to Predict Success Rates in Manual Tracking: Path Steering and Pursuit of 1D Moving Target
UIST '20· Human Pose & Activity Recognition +1
- 80%
Breathing Life Into Biomechanical User Models
UIST '22· Human Pose & Activity Recognition +1
- 67%
Computational Interaction: Theory and Practice
CHI '18· Computational Methods in HCI
- 67%
A Simulation Model of Intermittently Controlled Point-and-Click Behaviour
CHI '21· Eye Tracking & Gaze Interaction +2
- 67%
The I in Team: Mining Personal Social Interaction Routine with Topic Models from Long-Term Team Data
IUI '18· Human Pose & Activity Recognition +1
- 67%
MYND: Unsupervised Evaluation of Novel BCI Control Strategies on Consumer Hardware
UIST '20· Brain-Computer Interface (BCI) & Neurofeedback +1
- 60%
Computational Interaction: Theory and Practice
CHI '18· Computational Methods in HCI
- 60%
Computational Rationality as a Theory of Interaction
CHI '22· Computational Methods in HCI
Based on Jaccard similarity of research subtopics & professions (≥60%)