StyleFactory: Towards Better Style Alignment in Image Creation through Style-Strength-Based Control and Evaluation
Authors
Generative AI models have been widely used for image creation. However, generating images that are well-aligned with users' personal styles on aesthetic features (e.g., color and texture) can be challenging due to the poor style expression and interpretation between humans and models. Through a formative study, we observed that participants showed a clear subjective perception of the desired style and variations in its strength, which directly inspired us to develop style-strength-based control and evaluation. Building on this, we present StyleFactory, an interactive system that helps users achieve style alignment. Our interface enables users to rank images based on their strengths in the desired style and visualizes the strength distribution of other images in that style from the model's perspective. In this way, users can evaluate the understanding gap between themselves and the model, and define well-aligned personal styles for image creation through targeted iterations. Our technical evaluation and user study demonstrate that StyleFactory accurately generates images in specific styles, effectively facilitates style alignment in image creation workflow, stimulates creativity, and enhances the user experience in human-AI interactions.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 100%
Sketchforme: Composing Sketched Scenes from Text Descriptions for Interactive Applications
UIST '19· Generative AI (Text, Image, Music, Video) +2
- 83%
Examining the Text-to-Image Community of Practice: Why and How do People Prompt Generative AIs?
C&C '23· Generative AI (Text, Image, Music, Video) +1
- 67%
FlatMagic: Improving Flat Colorization through AI-driven Design for Digital Comic Professionals
CHI '22· Generative AI (Text, Image, Music, Video) +1
- 67%
GANravel: User-Driven Direction Disentanglement in Generative Adversarial Networks
CHI '23· Generative AI (Text, Image, Music, Video) +1
- 67%
Artinter: AI-powered Boundary Objects for Commissioning Visual Arts
DIS '23· Generative AI (Text, Image, Music, Video) +1
- 67%
When is a Tool a Tool? User Perceptions of System Agency in Human-AI Co-Creative Drawing
DIS '23· Generative AI (Text, Image, Music, Video) +1
- 67%
PromptPaint: Steering Text-to-Image Generation Through Paint Medium-like Interactions
UIST '23· Generative AI (Text, Image, Music, Video) +1
- 67%
DrawTalking: Building Interactive Worlds by Sketching and Speaking
UIST '24· AI-Assisted Creative Writing +1
Based on Jaccard similarity of research subtopics & professions (≥60%)