Diffuse Your Data Blues: Augmenting Low-Resource Datasets via User-Assisted Diffusion

Generative AI (Text, Image, Music, Video)3D Modeling & AnimationSoftware Engineers & DevelopersIndustrial Automation Engineers

Mixed reality applications in industrial contexts necessitate extensive and varied datasets for training object detection models, yet actual data gathering may be obstructed by logistical or cost issues. This study investigates the implementation of generative AI methods to work on this issue for mixed reality applications, with an emphasis on assembly and disassembly tasks. The novel objects found in industrial settings are difficult to describe using words, making text-based models less effective. In this study, a diffusion model is used to generate images by combining novel objects with various backgrounds. The backgrounds are selected where object detection in specific applications has been ineffective. This approach efficiently produces a diverse range of training samples. We compare three approaches: traditional augmentation methods, GAN-based augmentation, and Diffusion-based augmentation. Results show that the diffusion model significantly improved detection metrics. For instance, applying diffusion models to the dataset containing mechanical components of a pneumatic cylinder raised the $F1$ Score from $69.77$ to $84.21$ and the $mAP@50$ from $76.48$ to $88.77$, resulting in an $11\%$ increase in object detection performance, with a 67\% less dataset size compared to the traditional augmented dataset. The proposed image composition diffusion model and user-friendly interface further simplify dataset enrichment, proving effective for augmenting data and improving the robustness of detection models.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/iui/195777/2025

AdRecommended

Learn AI Coding at CodeNow

open_in_newOpen DOI Link
DOI: https://doi.org/10.1145/3708359.3712163
At a Glance

Paper Snapshot

fact_check
dataset
Source
IUI
calendar_month
Year
2025
emoji_events
Award
No award tagged
group
Authors
5 authors
sell
Subtopics
Generative AI (Text, Image, Music, Video), 3D Modeling & Animation
work
Professions
Software Engineers & Developers, Industrial Automation Engineers
article
Content Status
Abstract only
hub
Related Papers
0 related papers