PromptMap: An Alternative Interaction Style for AI-Based Image Generation
Authors
Recent technological advances popularized the use of image generation among the general public. Crafting effective prompts can, however, be difficult for novice users. To tackle this challenge, we developed PromptMap, a new interaction style for text-to-image AI that allows users to freely explore a vast collection of synthetic prompts through a map-like view with semantic zoom. PromptMap groups images visually by their semantic similarity, allowing users to discover relevant examples. We evaluated PromptMap in a between-subject online study (n=60) and a qualitative within-subject study (n=12). We found that PromptMap supported users in crafting prompts by providing them with examples. We also demonstrated the feasibility of using LLMs to create vast example collections. Our work contributes a new interaction style that supports users unfamiliar with prompting in achieving a satisfactory image output.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can users improve prompt writing efficiency and explore creative space when using text-to-image models?Category: Generative Image Creation and Editing ControlSimilar questionsarrow_forward
- Can a semantically structured example-browsing interface help users better control generation results and enhance creativity?Category: Generative Image Creation and Editing ControlSimilar questionsarrow_forward
- What are the specific advantages and challenges of large-scale synthetic datasets in generative AI interaction?Category: Generative Image Creation and Editing ControlSimilar questionsarrow_forward
Practical Problems
1- Novices struggle to generate appropriate prompts to control AI image generation results.Category: Generative Image Creation and Editing ControlSimilar questionsarrow_forward
- 60%
Dream Lens: Exploration and Visualization of Large-Scale Generative Design Datasets
CHI '18· Generative AI (Text, Image, Music, Video) +1
- 60%
UISketch: A Large-Scale Dataset of UI Element Sketches
CHI '21· Generative AI (Text, Image, Music, Video) +1
- 60%
PhotoScout: Synthesis-Powered Multi-Modal Image Search
CHI '24· Generative AI (Text, Image, Music, Video) +1
- 60%
Beyond the Ranked List: User-Driven Exploration and Diversification of Social Recommendation
IUI '18· Recommender System UX +1
Based on Jaccard similarity of research subtopics & professions (≥60%)