Preference-Conditioned Language-Guided Abstraction
Authors
Learning from demonstrations is a common paradigm for users to teach robots but is prone to spurious feature correlations. Recent work constructs state abstractions, i.e. visual representations containing only task-relevant features, from language as a way to perform more generalizable learning. However, these abstractions also depend on a user's preference for what matters in a task, which may be hard to describe or infeasible to exhaustively specify with language alone. How do we construct abstractions to capture these latent preferences? We observe that how humans behave reveals how they see the world. Our key insight is that differences in human behavior inform us that there are differences in preferences for how humans see the world, i.e. their state abstractions. In this work, we propose using language models (LMs) to query for those preferences directly given knowledge that a change in behavior has occurred. In our framework, we use the LM in two ways: first, given a text description of the task and knowledge of behavioral change between states, we query the LM for possible hidden preferences; second, given the most likely preference, we query the LM to construct the state abstraction. In this framework, the LM is also able to actively elicit human preferences when uncertain about its own estimate. We demonstrate our framework's ability to construct effective preference-conditioned abstractions in simulated experiments, a user study, as well as on a real Spot robot performing mobile manipulation tasks.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 60%
Effects of Communication Directionality and AI Agent Differences in Human-AI Interaction
CHI '21· Human-LLM Collaboration +1
- 60%
AI Knowledge: Improving AI Delegation through Human Enablement
CHI '23· Human-LLM Collaboration +1
- 60%
Comparing Sentence-Level Suggestions to Message-Level Suggestions in AI-Mediated Communication
CHI '23· Human-LLM Collaboration +1
- 60%
How Do Data Analysts Respond to AI Assistance? A Wizard-of-Oz Study
CHI '24· Human-LLM Collaboration +1
- 60%
VAL: Interactive Task Learning with GPT Dialog Parsing
CHI '24· Human-LLM Collaboration +1
- 60%
Towards Human-AI Deliberation: Design and Evaluation of LLM-Empowered Deliberative AI for AI-Assisted Decision-Making
CHI '25· Human-LLM Collaboration +1
- 60%
Need Help? Designing Proactive AI Assistants for Programming
CHI '25· Human-LLM Collaboration +1
- 60%
A Survey on Interactive Reinforcement Learning: Design Principles and Open Challenges
DIS '20· Human-LLM Collaboration +1
- 60%
Decision rule elicitation for domain adaptation
IUI '21· Human-LLM Collaboration +1
- 60%
Limitations of the LLM-as-a-Judge Approach for Evaluating LLM Outputs in Expert Knowledge Tasks
IUI '25· Human-LLM Collaboration +1
Based on Jaccard similarity of research subtopics & professions (≥60%)