Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning
Honorable MentionAuthors
We examine the problem of determining demonstration sufficiency: how can a robot self-assess whether it has received enough demonstrations from an expert to ensure a desired level of performance? To address this problem, we propose a novel self-assessment approach based on Bayesian inverse reinforcement learning and value-at-risk, enabling learning-from-demonstration ("LfD") robots to compute high-confidence bounds on their performance and use these bounds to determine when they have a sufficient number of demonstrations. We propose and evaluate two definitions of sufficiency: (1) normalized expected value difference, which measures regret with respect to the human's unobserved reward function, and (2) percent improvement over a baseline policy. We demonstrate how to formulate high-confidence bounds on both of these metrics. We evaluate our approach in simulation for both discrete and continuous state-space domains and illustrate the feasibility of developing a robotic system that can accurately evaluate demonstration sufficiency. We also show that the robot can utilize active learning in asking for demonstrations from specific states while evaluating demonstration sufficiency and that this results in fewer demos needed for the robot to still maintain high confidence in its policy. Finally, we demonstrate the viability of our approach via a user study.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 67%
TiiS: Artificial Intelligence for Modeling Complex Systems: Taming the Complexity of Expert Models to Improve Decision Making
IUI '22· AI-Assisted Decision-Making & Automation
- 60%
When is ML data good?: Valuing in Public Health Datafication
CHI '22· Teleoperated Driving +2
Based on Jaccard similarity of research subtopics & professions (≥60%)