A View on the Viewer: Gaze-Adaptive Captions for Videos
Honorable MentionAuthors
Subtitles play a crucial role in cross-lingual distribution of multimedia content and help communicate information where auditory content is not feasible (loud environments, hearing impairments, unknown languages). Established methods utilize text at the bottom of the screen, which may distract from the video. Alternative techniques place captions closer to related content (e.g., faces) but are not applicable to arbitrary videos such as documentations. Hence, we propose to leverage live gaze as indirect input method to adapt captions to individual viewing behavior. We implemented two gaze-adaptive methods and compared them in a user study (n=54) to traditional captions and audio-only videos. The results show that viewers with less experience with captions prefer our gaze-adaptive methods as they assist them in reading. Furthermore, gaze distributions resulting from our methods are closer to natural viewing behavior compared to the traditional approach. Based on these results, we provide design implications for gaze-adaptive captions.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 67%
EDITalk: Towards Designing Eyes-free Interactions for Mobile Word Processing
CHI '18· Voice User Interface (VUI) Design +1
- 67%
Adaptive Subtitles: Preferences and Trade-Offs in Real-Time Media Adaption
CHI '21· Voice Accessibility +1
- 67%
Acceptability of Speech and Silent Speech Input Methods in Private and Public
CHI '21· Voice User Interface (VUI) Design +1
- 67%
Designing for Speech Practice Systems: How Do User-Controlled Voice Manipulation and Model Speakers Impact Self-Perceptions of Voice?
CHI '22· Voice User Interface (VUI) Design +1
- 67%
Dynamik: Syntactically-Driven Dynamic Font Sizing for Emphasis of Key Information
IUI '25· Voice User Interface (VUI) Design +1
Based on Jaccard similarity of research subtopics & professions (≥60%)