RoboClean: Contextual Language Grounding for Human-Robot Interactions in Specialised Low-Resource Environments
Authors
Building effective voice interfaces for the instruction of service robots in specialised environments is difficult due to the local knowledge of workers, such as specific terminology for objects and space, leading to limited data to train language models (known as ‘low-resource’ domains) and challenges in language grounding. We present a language grounding study in which we a) elicit spoken natural language of context experts in situ through a Wizard of Oz study and compile a dataset, b) qualitatively examine linguistic properties of the resulting instructions to reveal referential categories and parameters employed to construct instructions in context. We discuss how our language grounding protocol may be applied to bootstrap a language model in its targeted use context. Our work contributes a linguistic understanding of robot instructions that can be applied by designers and researchers to develop spoken language understanding for human-robot interactions in specialised, low-resource environments.human-robot interaction, speech, conversational interfaces, lan- guage grounding, HRI, spoken language understanding, SLU
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can linguistic context in human–robot interaction be effectively semantically grounded in specific low-resource environments?Category: LLM Speech Dialogue and Communication RepairSimilar questionsarrow_forward
- Can analysis and classification of natural language corpora significantly improve robots' understanding of specialized commands?Category: LLM Speech Dialogue and Communication RepairSimilar questionsarrow_forward
- How well does language data collected through Wizard-of-Oz (WOz) methods adapt to low-resource environments?Category: LLM Speech Dialogue and Communication RepairSimilar questionsarrow_forward
Practical Problems
1- In specific low-resource environments, robots struggle to understand specialized linguistic commands.Category: LLM Speech Dialogue and Communication RepairSimilar questionsarrow_forward
- 80%
Human Speakers Help Machine Listeners To account For Visual Asymmetries in Dialogue
CUI '23· Voice User Interface (VUI) Design +1
- 67%
Tea, Earl Grey, Hot: Designing Speech Interactions from the Imagined Ideal of Star Trek
CHI '21· Voice User Interface (VUI) Design +2
- 67%
Enabling Conversational Interaction with Mobile UI using Large Language Models
CHI '23· Voice User Interface (VUI) Design +1
- 67%
ONYX: Assisting Users in Teaching Natural Language Interfaces Through Multi-Modal Interactive Task Learning
CHI '23· Voice User Interface (VUI) Design +2
- 67%
Discovering Natural Language Commands in Multimodal Interfaces
IUI '19· Voice User Interface (VUI) Design +2
- 67%
Screen2Words: Automatic Mobile UI Summarization with Multimodal Learning
UIST '21· Voice User Interface (VUI) Design +1
- 60%
Keep it Short: A Comparison of Voice Assistants' Response Behavior
CHI '22· Voice User Interface (VUI) Design +1
- 60%
What’s The Talk on VUI Guidelines? A Meta-Analysis of Guidelines for Voice User Interface Design
CUI '23· Voice User Interface (VUI) Design
- 60%
VOICON: Geometric Motion-Based Visual Feedback in Voice User Interface
DIS '24· Voice User Interface (VUI) Design +1
- 60%
Cells, Generators, and Lenses: Design Framework for Object-Oriented Interaction with Large Language Models
UIST '23· Human-LLM Collaboration
Based on Jaccard similarity of research subtopics & professions (≥60%)