Improving Subject Representation in AI Generated Art: Design Guidelines for Using Image Prompts with Text-to-Image Generative Models

Generative AI (Text, Image, Music, Video)AI-Assisted Creative WritingGraphic Design & Typography ToolsFilm & Animation ProducersUI/UX DesignersVisual Artists & Designers

Advances in text-to-image generative models have made it easier for people to create art by just prompting models with text. However, creating through text leaves users with limited control over the final composition or the way the subject is represented. A potential solution is to use image prompts alongside text prompts to condition the model. To better understand how and when image prompts can improve subject representation in generations, we conduct an annotation experiment to quantify their effect on generations of abstract, concrete plural, and concrete singular subjects. We find that initial images improved subject representation across all subject types, with the most noticeable improvement in concrete singular subjects. In an analysis of different types of initial images, we find that icons and photos produced high quality generations of different aesthetics. We conclude with design guidelines for how initial images can improve subject representation in AI art.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/cc/83009/2022

AdRecommended

Learn AI Coding at CodeNow

open_in_newOpen DOI Link
DOI: https://doi.org/10.1145/3527927.3532792
At a Glance

Paper Snapshot

fact_check
dataset
Source
C&C
calendar_month
Year
2022
emoji_events
Award
No award tagged
group
Authors
3 authors
sell
Subtopics
Generative AI (Text, Image, Music, Video), AI-Assisted Creative Writing, Graphic Design & Typography Tools
work
Professions
Film & Animation Producers, UI/UX Designers, Visual Artists & Designers
article
Content Status
Abstract only
hub
Related Papers
8 related papers