Online Behavior Modification for Expressive User Control of RL-Trained Robots
Authors
Reinforcement Learning (RL) is an effective method for robots to learn tasks. However, in typical RL, end-users have little to no control over how the robot does the task after the robot has been deployed. To address this, we introduce the idea of online behavior modification, a paradigm in which users have control over behavior features of a robot in real-time as it autonomously completes a task using an RL-trained policy. To show the value of this user-centered formulation for human-robot interaction, we present a behavior-diversity–based algorithm, Adjustable Control Of RL Dynamics (ACORD), and demonstrate its applicability to online behavior modification in simulation and a user study. In the study (n=23), users adjust the style of paintings as a robot traces a shape autonomously. We compare ACORD to RL and Shared Autonomy (SA), and show ACORD affords user-preferred levels of control and expression, comparable to SA, but with the potential for autonomous execution and robustness of RL. The code for this paper is available at anon.url
Research Questions / Practical Problems
Question signals indexed for this paper.
- 100%
Aligning Human and Robot Representations
HRI '24· AI-Assisted Decision-Making & Automation +1
- 80%
REX: Designing User-centered Repair and Explanations to Address Robot Failures
DIS '24· Explainable AI (XAI) +2
- 80%
Enhancing Safety in Learning from Demonstration Algorithms via Control Barrier Function Shielding
HRI '24· AI-Assisted Decision-Making & Automation +1
- 67%
"Should I Rely on You or the AI?" Leaders' Trust and Perceptions in Mixed Human-AI Teams
CHI '26· Human-Robot Collaboration (HRC) +2
- 67%
Friend, Foe, or Bot? Exploring Intergroup Dynamics in Hybrid Human-Bot Teams
CHI '26· Human-Robot Collaboration (HRC) +2
- 60%
Effects of Communication Directionality and AI Agent Differences in Human-AI Interaction
CHI '21· Human-LLM Collaboration +1
- 60%
Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performance
CHI '21· Explainable AI (XAI) +1
- 60%
Two Heads Are Better Than One: A Dimension Space for Unifying Human and Artificial Intelligence in Shared Control
CHI '22· AI-Assisted Decision-Making & Automation +1
- 60%
Model Sketching: Centering Concepts in Early-Stage Machine Learning Model Design
CHI '23· AI-Assisted Decision-Making & Automation +1
- 60%
Matching Mind and Method: Augmented Decision-Making with Digital Companions based on Regulatory Mode Theory
CHI '23· AI-Assisted Decision-Making & Automation
Based on Jaccard similarity of research subtopics & professions (≥60%)