DoThisHere: Using multi-modal interaction to support cross-application tasks on mobile devices
Authors
Many computing tasks, such as comparison shopping, twofactor authentication, and checking movie reviews, require using multiple apps together. On large screens, “windows, icons, menus, pointer” (WIMP) graphical user interfaces (GUIs) support easy sharing of content and context between multiple apps. So, it is easy to see the content from one application and write something relevant in another application, such as looking at the map around a place and typing walking instructions into an email. However, although today’s smartphones also use GUIs, they have small screens and limited windowing support, making it hard to switch contexts and exchange data between apps. We introduce DoThisHere, a multimodal interaction technique that streamlines cross-app tasks and reduces the burden thesetasks impose on users. Users can use voice to refer to information or app features that are off-screen and touch to specify where the relevant information should be inserted or is displayed. With DoThisHere, users can access information from or carry information to other apps with less context switching. We conducted a survey to find out what cross-app tasks people are performing or wish to perform on their smartphones. Among the 125 tasks that we collected from 75 participants, we found that 59 of these tasks are not well supported currently. DoThisHere is helpful in completing 95% of these unsupported tasks. A user study, where users are shown the list of supported voice commands when performing a representative sample of such tasks, suggests that DoThisHere may reduce expert users’ cognitive load; the Query action, in particular, can help users reduce task completion time.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 60%
Challenges and Opportunities for Technology-Supported Activity Reporting in the Workplace
CHI '18· Knowledge Management & Team Awareness +1
- 60%
ActiveErgo: Automatic and Personalized Ergonomics using Self-actuating Furniture
CHI '18· Full-Body Interaction & Embodied Input +1
- 60%
The Impact of Word, Multiple Word, and Sentence Input on Virtual Keyboard Decoding Performance
CHI '18· Voice User Interface (VUI) Design +1
- 60%
Exploring New File Metaphors for a Networked World through the File Biography
CHI '18· Knowledge Worker Tools & Workflows +1
- 60%
Doppio: Tracking UI Flows and Code Changes for App Development
CHI '18· Knowledge Worker Tools & Workflows +1
- 60%
Pointing at a Distance with Everyday Smart Devices
CHI '18· Full-Body Interaction & Embodied Input +1
- 60%
Using Visual Histories to Reconstruct the Mental Context of Suspended Activities
CHI '18· Knowledge Worker Tools & Workflows +1
- 60%
Leveraging Community-Generated Videos and Command Logs to Classify and Recommend Software Workflows
CHI '18· Crowdsourcing Task Design & Quality Control +1
- 60%
HotStrokes: Word-Gesture Shortcuts on a Trackpad
CHI '19· Hand Gesture Recognition +1
- 60%
GazeConduits: Calibration-Free Cross-Device Collaboration through Gaze and Touch
CHI '20· Eye Tracking & Gaze Interaction +1
Based on Jaccard similarity of research subtopics & professions (≥60%)