GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment Design
Authors
Environment designers in the entertainment industry create imaginative 2D and 3D scenes for games, films, and television, requiring both fine-grained control of specific details and consistent global coherence. Designers have increasingly integrated generative AI into their workflows, often relying on large language models (LLMs) to expand user prompts for text-to-image generation, then iteratively refining those prompts and applying inpainting. However, our formative study with 10 designers surfaced two key challenges: (1) the lengthy LLM-generated prompts make it difficult to understand and isolate the keywords that must be revised for specific visual elements; and (2) while inpainting supports localized edits, it can struggle with global consistency and correctness. Based on these insights, we present GenTune, an approach that enhances human–AI collaboration by clarifying how AI-generated prompts map to image content. Our GenTune system lets designers select any element in a generated image, trace it back to the corresponding prompt labels, and revise those labels to guide precise yet globally consistent image refinement. In a summative study with 20 designers, GenTune significantly improved prompt-image comprehension, refinement quality and efficiency, and overall satisfaction (all p < .01) compared to current practice. A follow-up field study with two studios further demonstrated its effectiveness in real-world settings.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 100%
Jigsaw: Supporting Designers to Prototype Multimodal Applications by Chaining AI Foundation Models
CHI '24· Generative AI (Text, Image, Music, Video) +2
- 80%
Exploring Challenges and Opportunities to Support Designers in Learning to Co-create with AI-based Manufacturing Design Tools
CHI '23· Generative AI (Text, Image, Music, Video) +1
- 80%
AI-Assisted Causal Pathway Diagram for Human-Centered Design
CHI '24· Generative AI (Text, Image, Music, Video) +1
- 80%
I-Card: A Generative AI-Supported Intelligent Design Method Card Deck
CHI '25· Generative AI (Text, Image, Music, Video) +1
- 80%
Design Ideation with AI - Sketching, Thinking and Talking with Generative Machine Learning Models
DIS '23· Generative AI (Text, Image, Music, Video) +1
- 80%
PromptInfuser: How Tightly Coupling AI and UI Design Impacts Designers’ Workflows
DIS '24· Human-LLM Collaboration +1
- 67%
May AI? Design Ideation with Cooperative Contextual Bandits
CHI '19· Generative AI (Text, Image, Music, Video) +2
- 67%
ICONATE: Automatic Compound Icon Generation and Ideation
CHI '20· Generative AI (Text, Image, Music, Video) +2
- 67%
PlantoGraphy: Incorporating Iterative Design Process into Generative Artificial Intelligence for Landscape Rendering
CHI '24· Generative AI (Text, Image, Music, Video) +2
- 67%
IEDS: Exploring an Intelli-Embodied Design Space Combining Designer, AR, and GAI to Support Industrial Conceptual Design
CHI '25· AR Navigation & Context Awareness +2
Based on Jaccard similarity of research subtopics & professions (≥60%)