Discovering Natural Language Commands in Multimodal Interfaces
Authors
Discovering what to say and how to say it remains a challenge for users of multimodal interfaces supporting speech input. Users end up "guessing" commands that a system might support, often leading to interpretation errors and frustration. One solution to this problem is to display contextually relevant command examples as users interact with a system. The challenge, however, is deciding when, how, and which examples to recommend. In this work, we describe an approach for generating and ranking natural language command examples in multimodal interfaces. We demonstrate the approach using a prototype touch- and speech-based image editing tool. We experiment with augmentations of the UI to understand when and how to present command examples. Through an online user study, we evaluate these alternatives and find that in-situ command suggestions promote discovery and encourage the use of speech input.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 67%
Designing Voice Interfaces: Back to the (Curriculum) Basics
CHI '20· Voice User Interface (VUI) Design +2
- 67%
Tea, Earl Grey, Hot: Designing Speech Interactions from the Imagined Ideal of Star Trek
CHI '21· Voice User Interface (VUI) Design +2
- 67%
RoboClean: Contextual Language Grounding for Human-Robot Interactions in Specialised Low-Resource Environments
CUI '23· Voice User Interface (VUI) Design +2
- 60%
Keep it Short: A Comparison of Voice Assistants' Response Behavior
CHI '22· Voice User Interface (VUI) Design +1
- 60%
What’s The Talk on VUI Guidelines? A Meta-Analysis of Guidelines for Voice User Interface Design
CUI '23· Voice User Interface (VUI) Design
- 60%
VOICON: Geometric Motion-Based Visual Feedback in Voice User Interface
DIS '24· Voice User Interface (VUI) Design +1
Based on Jaccard similarity of research subtopics & professions (≥60%)