Semantic See-through Goggles: Wearing Linguistic Virtual Reality in (Artificial Intelligence)
Authors
When language is used as a medium for sensory information—as when one describes a scene in words—a kind of virtual reality emerges: realities projected into the same sentence become virtual/equivalent. We call this Linguistic VR (LVR). LVR is constituted by codes shaped chiefly by majority cultural norms. In today’s world—where AI, largely built on linguistic data and processes, is deeply entangled with the everyday mediation of sensory information—it is necessary to critically re-examine the virtuality of this VR. We propose Semantic See-through Goggles, a system that makes manifest, as a first-person experience, the LVR latent in the linguistic mediation of scenes, enabling intuitive understanding and analysis of its properties and issues. The system inserts a serial image-to-text and text-to-image transformation pipeline between the camera and head-mounted display (HMD) of video see-through goggles: the live view is converted into a single line of text, then re-generated as an image, and only this mediated image reaches the wearer’s eyes. We built a prototype and validated its basic properties, followed by a qualitative analysis of wearer experiences. The results suggest that this method enables subjective experience—and thus understanding—of the transformations of environmental information, and of the attendant issues, induced by the linguistic mediation of vision. We also obtained preliminary insights into both AI-driven linguistic mediation of sensory information and problems intrinsic to linguistic mediation itself.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 86%
The Metacognitive Demands and Opportunities of Generative AI
CHI '24· Generative AI (Text, Image, Music, Video) +2
- 71%
Think Together and Work Better: Combining Humans' and LLMs' Think-Aloud Outcomes for Effective Text Evaluation
CHI '25· Generative AI (Text, Image, Music, Video) +2
- 63%
“I’m happy even though it’s not real”: GenAI Photo Editing as a Remembering Experience
CHI '26· Generative AI (Text, Image, Music, Video) +2
- 63%
Personal Validation Effect in LLMs: Positive AI Responses Bias Perceptions of Validity, Reliability, Personalization, and Usefulness of Fictitious Predictions
CHI '26· Human-LLM Collaboration +2
- 63%
Tell Me What I Missed: Interacting with GPT during Recalling of One-Time Witnessed Events
CHI '26· Human-LLM Collaboration +2
- 63%
A Survey of Collaborative Reinforcement Learning: Interactive Methods and Design Patterns
DIS '21· Human-LLM Collaboration +2
- 63%
PersonaFlow: Designing LLM-Simulated Expert Perspectives for Enhanced Research Ideation
DIS '25· Generative AI (Text, Image, Music, Video) +2
- 63%
Exploring the Innovation Opportunities for Pre-trained Models
DIS '25· Generative AI (Text, Image, Music, Video) +2
- 63%
Who Needs What Explanation? How User Traits Affect Explanation Effectiveness in AI-Assisted Decision-Making
IUI '26· AI-Assisted Decision-Making & Automation +2
Based on Jaccard similarity of research subtopics & professions (≥60%)