Lights, Camera, Autonomy! Generative Video Fusion for Realistic In Situ Evaluation of Autonomous Drones
Authors
Autonomous service robots and drones are increasingly deployed in urban spaces, yet evaluating their designs in situ is constrained by permits, logistics, and the cost of repeated user studies. We contribute SHOWTIME, a generative video-fusion framework that synthesizes realistic, controllable evaluation clips by combining structured prompts, encoding navigation behavior, proxemic policies, and task context with real urban scenes. Across two studies, participants rated SHOWTIME’s generative videos on par with physical recordings for realism, behavioral plausibility, ecological appropriateness, and comfort, and could not reliably distinguish real from generated footage; perceptual differences were driven by environmental structure and crowd density rather than modality. A follow-up experiment showed that generative prototyping elicits robust perceptual contrasts across design variants: communicative signaling (lights, functional cues) increases acceptability and comfort, whereas erratic motion reduces them. The pipeline achieves over a 10× reduction in per-clip time compared to physical trials, enabling rapid, human-in-the-loop iteration on behaviors and UI cues. Together, these results demonstrate human-level perceptual fidelity and practical viability, positioning SHOWTIME as a cost-effective, scalable method for refining autonomous systems and their interfaces in real-world urban contexts.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 63%
"What do I do now?": Spontaneous Human Responses to Robot Effectiveness and Efficiency Malfunctions in Collaborative Robotics
CHI '26· Human-Robot Collaboration (HRC) +2
- 63%
PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
HRI '24· Generative AI (Text, Image, Music, Video) +2
- 63%
Guidance Source Matters: How Guidance from AI, Expert, or a Group of Analysts Impacts Visual Data Preparation and Analysis
IUI '25· Generative AI (Text, Image, Music, Video) +2
Based on Jaccard similarity of research subtopics & professions (≥60%)