Identifying Speech Input Errors Through Audio-Only Interaction

Voice AccessibilityContent Creators (YouTubers, Podcasters)

Speech has become an increasingly common means of text input, from smartphones and smartwatches to voice-based intelligent personal assistants. However, reviewing the recognized text to identify and correct errors is a challenge when no visual feedback is available. In this paper, we first quantify and describe the speech recognition errors that users are prone to miss, and investigate how to better support this error identification task by manipulating pauses between words, speech rate, and speech repetition. To achieve these goals, we conducted a series of four studies. Study 1, an in-lab study, showed that participants missed identifying over 50% of speech recognition errors when listening to audio output of the recognized text. Building on this result, Studies 2 to 4 were conducted using an online crowdsourcing platform and showed that adding a pause between words improves error identification compared to no pause, the ability to identify errors degrades with higher speech rates (300 WPM), and repeating the speech output does not improve error identification. We derive implications for the design of audio-only speech dictation.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/chi/3809/2018

AdRecommended

Learn AI Coding at CodeNow

At a Glance

Paper Snapshot

fact_check
dataset
Source
CHI
calendar_month
Year
2018
emoji_events
Award
No award tagged
group
Authors
2 authors
sell
Subtopics
Voice Accessibility
work
Professions
Content Creators (YouTubers, Podcasters)
article
Content Status
Abstract only
hub
Related Papers
0 related papers