Enhancing Visitor Engagement in Interactive Art Exhibitions with Visual-Enhanced Conversational Agents
Authors
Conversational agents in art exhibitions can enhance user engagement and understanding of artworks by providing contextual information, especially through voice interactions. However, creating a deeper personal connection with art - which often requires direct aesthetic and visual experiences - remains a challenge. This paper examines how integrating visual perception into conversational agents can enhance alignment with visitors' artistic interpretations, thereby fostering deeper engagement with interactive art exhibitions. We introduce a voice-based conversational agent enhanced with visual capabilities via a multimodal large language model (MLLM), allowing the agent to perceive and discuss artworks in real-time with visitors. The system utilizes a simplified Retrieval-Augmented Generation (RAG) architecture, which collects voice inputs, retrieves relevant information from a domain knowledge graph, and uses the LLM to generate conversational responses, which are then converted into voice outputs. A user study with 36 participants, divided into two groups, was conducted to compare the enhanced system with a baseline system that lacked visual input. Results show that the visually enhanced system significantly improved visitor engagement and satisfaction. Content analysis of the conversational transcripts further revealed a wider range of conversational topics, deeper visitor perceptions, and the agent's ability to provide more nuanced, visually-related discussions.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can traditional voice interaction assistants use visual enhancement to improve users' deep engagement with art in exhibitions?Category: Voice Assistant General Design and User ExperienceSimilar questionsarrow_forward
- How can multimodal dialogue assistants dynamically analyze visual elements in artworks and generate meaningful discussion content?Category: Voice Assistant General Design and User ExperienceSimilar questionsarrow_forward
- How can knowledge graphs and multimodal LLMs improve content relevance and interaction experience in art exhibition dialogue assistants?Category: Voice Assistant General Design and User ExperienceSimilar questionsarrow_forward
Practical Problems
1- Voice assistants in museums struggle to help visitors deeply understand artworks.Category: Voice Assistant General Design and User ExperienceSimilar questionsarrow_forward
No related papers with ≥60% similarity
Based on Jaccard similarity of research subtopics & professions (≥60%)