Human Speakers Help Machine Listeners To account For Visual Asymmetries in Dialogue
Authors
Human-machine dialogue (HMD) research debates the degree to which language production in this context is egocentric or allocentric. That is, the degree to which a person might take a machine’s perspective into account. Our study aims to identify whether users produce allocentric or egocentric language within speech-based HMD when there is asymmetry in the information available to both partners. Through an adapted referential communication task, we manipulated the presence or absence of visual distractors and occlusions, similarly to previous referential tasks used in psycholinguistic research. Results show that people are sensitive to the presence of distractors and occlusions and tend to produce more informative expressions to help machine partners account for the visual asymmetries. We discuss the fndings on how allocentric production in HMD is explained by how the division of labour manifests in spoken HMD. The fndings further our understanding of the language production mechanisms in HMD.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- In dialogues with visual obstacles such as distractors or occlusions, how do humans adjust language to help machines understand their referential perspective?Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
- How do visual ambiguity (e.g., distractors) and uncertainty (e.g., occlusions) affect the granularity of human language?Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
- Under the same visual conditions, do behavioral patterns of language expression differ between human-human and human-machine dialogue?Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
Practical Problems
1- Smart voice assistants struggle to understand users' language instructions under visually obstructed conditions.Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
- 80%
Enabling Conversational Interaction with Mobile UI using Large Language Models
CHI '23· Voice User Interface (VUI) Design +1
- 80%
ONYX: Assisting Users in Teaching Natural Language Interfaces Through Multi-Modal Interactive Task Learning
CHI '23· Voice User Interface (VUI) Design +2
- 80%
RoboClean: Contextual Language Grounding for Human-Robot Interactions in Specialised Low-Resource Environments
CUI '23· Voice User Interface (VUI) Design +2
- 80%
Screen2Words: Automatic Mobile UI Summarization with Multimodal Learning
UIST '21· Voice User Interface (VUI) Design +1
- 75%
What’s The Talk on VUI Guidelines? A Meta-Analysis of Guidelines for Voice User Interface Design
CUI '23· Voice User Interface (VUI) Design
- 75%
Cells, Generators, and Lenses: Design Framework for Object-Oriented Interaction with Large Language Models
UIST '23· Human-LLM Collaboration
- 67%
ReactGenie: A Development Framework for Complex Multimodal Interactions Using Large Language Models
CHI '24· Voice User Interface (VUI) Design +2
- 67%
Talk to the Hand: an LLM-powered Chatbot with Visual Pointer as Proactive Companion for On-Screen Tasks
CHI '25· Voice User Interface (VUI) Design +2
- 67%
VoiceAlign: A Shimming Layer for Enhancing the Usability of Legacy Voice User Interface Systems
IUI '26· Voice User Interface (VUI) Design +2
- 67%
"Same Voice, Different Language": An Exploration of Voice-Cloned Translation to Support Non-Native Speakers in Online Meetings
IUI '26· Multilingual & Cross-Cultural Voice Interaction +2
Based on Jaccard similarity of research subtopics & professions (≥60%)