HRI '241st AuthorAI-Assisted Decision-Making & Automation +1
Modeling Variation in Human Feedback with User Inputs: An Exploratory Methodology
To expedite the development process of interactive reinforcement learning (IntRL) algorithms, prior work often uses perfect oracles as simulated human teachers to furnish feedback signals. Those oracles typically derive from ground-truth knowledge or optimal policies, and provide dense and error-free feedback to a rob…