LabelAR: A Spatial Guidance Interface for Fast Computer Vision Image Collection
Authors
Computer vision is applied in an ever expanding range of applications, many of which require custom training data to perform well. We present a novel interface for rapid collection and labeling of training images to improve computer vision based object detectors. LabelAR leverages the spatial tracking capabilities of an AR-enabled camera, allowing users to place persistent bounding volumes that stay centered on real-world objects. The interface then guides the user to move the camera to cover a wide variety of viewpoints. We eliminate the need for post-hoc manual labeling of images by automatically projecting 2D bounding boxes around objects in the images as they are captured from AR-marked viewpoints. Across 12 users, LabelAR significantly outperforms existing approaches in terms of the trade-off between model performance and collection time.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 67%
CARING-AI: Towards Authoring Context-aware Augmented Reality INstruction through Generative Artificial Intelligence
CHI '25· AR Navigation & Context Awareness +2
- 60%
Discovering the Syntax and Strategies of Natural Language Programming with Generative Language Models
CHI '22· Generative AI (Text, Image, Music, Video) +1
- 60%
"What It Wants Me To Say": Bridging the Abstraction Gap Between End-User Programmers and Code-Generating Large Language Models
CHI '23· Generative AI (Text, Image, Music, Video) +1
- 60%
D-Twins: Your Digital Twin Designed for Real-Time Boredom Intervention
CHI '25· Generative AI (Text, Image, Music, Video) +1
- 60%
Augmented Silkscreen: Designing AR Interactions for Debugging Printed Circuit Boards
DIS '21· AR Navigation & Context Awareness +1
- 60%
From Discovery to Adoption: Understanding the ML Practitioners' Interpretability Journey
DIS '23· Generative AI (Text, Image, Music, Video) +1
- 60%
Take It, Leave It, or Fix It: Measuring Productivity and Trust in Human-AI Collaboration
IUI '24· Generative AI (Text, Image, Music, Video) +1
- 60%
From Interaction to Impact: Towards Safer AI Agent Through Understanding and Evaluating Mobile UI Operation Impacts
IUI '25· Generative AI (Text, Image, Music, Video) +1
Based on Jaccard similarity of research subtopics & professions (≥60%)