SwipeGANSpace: Swipe-to-Compare Image Generation via Efficient Latent Space Exploration
Authors
Generating preferred images using generative adversarial networks (GANs) is challenging owing to the high-dimensional nature of latent space. In this study, we propose a novel approach that uses simple user-swipe interactions to generate preferred images for users. To effectively explore the latent space with only swipe interactions, we apply principal component analysis to the latent space of the StyleGAN, creating meaningful subspaces. We use a multi-armed bandit algorithm to decide the dimensions to explore, focusing on the preferences of the user. Experiments show that our method is more efficient in generating preferred images than the baseline methods. Furthermore, changes in preferred images during image generation or the display of entirely different image styles were observed to provide new inspirations, subsequently altering user preferences. This highlights the dynamic nature of user preferences, which our proposed approach recognizes and enhances.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can gesture interaction optimize StyleGAN latent space exploration to generate user-preferred images?Category: Gesture Sensing, Recognition Algorithms, and Sensor TechnologiesSimilar questionsarrow_forward
- After reducing latent space dimensionality, how can key generative directions for user preferences be efficiently located?Category: Gesture Sensing, Recognition Algorithms, and Sensor TechnologiesSimilar questionsarrow_forward
- Can swipe-gesture-based interaction effectively adapt to users' dynamically changing preferences?Category: Gesture Sensing, Recognition Algorithms, and Sensor TechnologiesSimilar questionsarrow_forward
Practical Problems
1- Users find it complex and difficult to express needs when generating preferred images with GANs on small-screen devices.Category: Gesture Sensing, Recognition Algorithms, and Sensor TechnologiesSimilar questionsarrow_forward
- 100%
Design Guidelines for Prompt Engineering Text-to-Image Generative Models
CHI '22· Generative AI (Text, Image, Music, Video)
- 100%
FusAIn: Composing Generative AI Visual Prompts Using Pen-based Interaction
CHI '25· Generative AI (Text, Image, Music, Video)
- 75%
FlatMagic: Improving Flat Colorization through AI-driven Design for Digital Comic Professionals
CHI '22· Generative AI (Text, Image, Music, Video) +1
- 75%
GANravel: User-Driven Direction Disentanglement in Generative Adversarial Networks
CHI '23· Generative AI (Text, Image, Music, Video) +1
- 75%
Personalizing Products with Stylized Head Portraits for Self-Expression
CHI '24· Generative AI (Text, Image, Music, Video) +1
- 75%
GenColor: Generative Color-Concept Association in Visual Design
CHI '25· Generative AI (Text, Image, Music, Video) +1
- 75%
When is a Tool a Tool? User Perceptions of System Agency in Human-AI Co-Creative Drawing
DIS '23· Generative AI (Text, Image, Music, Video) +1
- 75%
GenFrame – Embedding Generative AI Into Interactive Artifacts
DIS '24· Generative AI (Text, Image, Music, Video) +1
- 75%
Continuous and Gradual Style Changes of Graphic Designs with Generative Model
IUI '21· Generative AI (Text, Image, Music, Video) +1
- 60%
"I don't want to feel like I'm working in a 1960s factory": The Practitioner Perspective on Creativity Support Tool Adoption
CHI '22· Generative AI (Text, Image, Music, Video) +1
Based on Jaccard similarity of research subtopics & professions (≥60%)