Improving Subject Representation in AI Generated Art: Design Guidelines for Using Image Prompts with Text-to-Image Generative Models
Advances in text-to-image generative models have made it easier for people to create art by just prompting models with text. However, creating through text leaves users with limited control over the final composition or the way the subject is represented. A potential solution is to use image prompts alongside text prompts to condition the model. To better understand how and when image prompts can improve subject representation in generations, we conduct an annotation experiment to quantify their effect on generations of abstract, concrete plural, and concrete singular subjects. We find that initial images improved subject representation across all subject types, with the most noticeable improvement in concrete singular subjects. In an analysis of different types of initial images, we find that icons and photos produced high quality generations of different aesthetics. We conclude with design guidelines for how initial images can improve subject representation in AI art.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 71%
Prompting for Discovery: Flexible Sense-Making for AI Art-Making with Dreamsheets
CHI '24· Generative AI (Text, Image, Music, Video) +2
- 71%
Creative Blends of Visual Concepts
CHI '25· Generative AI (Text, Image, Music, Video) +2
- 67%
FlatMagic: Improving Flat Colorization through AI-driven Design for Digital Comic Professionals
CHI '22· Generative AI (Text, Image, Music, Video) +1
- 67%
GenColor: Generative Color-Concept Association in Visual Design
CHI '25· Generative AI (Text, Image, Music, Video) +1
- 67%
Generative AI in Documentary Photography: Exploring Opportunities and Challenges for Visual Storytelling
CHI '25· Generative AI (Text, Image, Music, Video) +1
- 67%
When is a Tool a Tool? User Perceptions of System Agency in Human-AI Co-Creative Drawing
DIS '23· Generative AI (Text, Image, Music, Video) +1
- 67%
Continuous and Gradual Style Changes of Graphic Designs with Generative Model
IUI '21· Generative AI (Text, Image, Music, Video) +1
- 63%
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
UIST '25· Generative AI (Text, Image, Music, Video) +1
Based on Jaccard similarity of research subtopics & professions (≥60%)