MemoVis: A GenAI-Powered Tool for Creating Companion Reference Images for 3D Design Feedback
Authors
Providing asynchronous feedback is a critical step in the 3D design workflow. A common approach to providing feedback is to pair textual comments with companion reference images, which helps illustrate the gist of text. Ideally, feedback providers should possess 3D and image editing skills to create reference images that can effectively describe what they have in mind. However, they often lack such skills, so they have to resort to sketches or online images which might not match well with the current 3D design. To address this, we introduce MemoVis, a text editor interface that assists feedback providers in creating reference images with generative AI driven by the feedback comments. First, a novel real-time viewpoint suggestion feature, based on a vision-language foundation model, helps feedback providers anchor a comment with a camera viewpoint. Second, given a camera viewpoint, we introduce three types of image modifiers, based on pre-trained 2D generative models, to turn a text comment into an updated version of the 3D scene from that viewpoint. We conducted a within-subjects study with 14 feedback providers, demonstrating the effectiveness of MemoVis. The quality and explicitness of the companion images were evaluated by another eight participants with prior 3D design experience.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 83%
Scaling Creative Inspiration with Fine-Grained Functional Aspects of Product Ideas
CHI '22· Generative AI (Text, Image, Music, Video) +1
- 71%
GenPara: Enhancing the 3D Design Editing Process by Inferring Users' Regions of Interest with Text-Conditional Shape Parameters
CHI '25· Generative AI (Text, Image, Music, Video) +2
- 71%
Exploring the Potential of Metacognitive Support Agents for Human-AI Co-Creation
DIS '25· Generative AI (Text, Image, Music, Video) +2
- 67%
The Impact of Sketch-guided vs. Prompt-guided 3D Generative AIs on the Design Exploration Process
CHI '24· Generative AI (Text, Image, Music, Video) +1
- 67%
Exploring Optimal Combinations: The Impact of Sequential Multimodal Inspirational Stimuli in Design Concepts on Creativity
DIS '24· Generative AI (Text, Image, Music, Video) +1
- 67%
CrossGAI: A Cross-Device Generative AI Framework for Collaborative Fashion Design
UbiComp '24· Generative AI (Text, Image, Music, Video) +1
- 63%
"I Just Need GPT to Refine My Prompts”: Rethinking Onboarding and Help-Seeking with Generative 3D Modelling Tools
CHI '26· Generative AI (Text, Image, Music, Video) +3
Based on Jaccard similarity of research subtopics & professions (≥60%)