Audible Panorama: Automatic Spatial Audio Generation for Panorama Imagery

Head-Up Display (HUD) & Advanced Driver Assistance Systems (ADAS)360° Video & Panoramic ContentGenerative AI (Text, Image, Music, Video)

As 360 deg cameras and virtual reality headsets become more popular, panorama images have become increasingly ubiquitous. While sounds are essential in delivering immersive and interactive user experiences, most panorama images, however, do not come with native audio. In this paper, we propose an automatic algorithm to augment static panorama images through realistic audio assignment. We accomplish this goal through object detection, scene classification, object depth estimation, and audio source placement. We built an audio file database composed of over $500$ audio files to facilitate this process. We designed and conducted a user study to verify the efficacy of various components in our pipeline. We run our method on a large variety of panorama images of indoor and outdoor scenes. By analyzing the statistics, we learned the relative importance of these components, which can be used in prioritizing for power-sensitive time-critical tasks like mobile augmented reality (AR) applications.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/chi/5733/2019

AdRecommended

Learn AI Coding at CodeNow

At a Glance

Paper Snapshot

fact_check
dataset
Source
CHI
calendar_month
Year
2019
emoji_events
Award
No award tagged
group
Authors
4 authors
sell
Subtopics
Head-Up Display (HUD) & Advanced Driver Assistance Systems (ADAS), 360° Video & Panoramic Content, Generative AI (Text, Image, Music, Video)
work
Professions
article
Content Status
Abstract only
hub
Related Papers
0 related papers