MIMOSA: Human-AI Co-Creation of Computational Spatial Audio Effects on Videos

Generative AI (Text, Image, Music, Video)Music Composition & Sound Design ToolsCreative Collaboration & Feedback SystemsContent Creators (YouTubers, Podcasters)Musicians, DJs & Sound DesignersFilm & Animation Producers

Spatial audio offers more immersive video consumption experiences to viewers; however, creating and editing spatial audio often expensive and requires specialized equipment and skills, posing a high barrier for amateur video creators. We present MIMOSA, a human-AI co-creation tool that enables amateur users to computationally generate and manipulate spatial audio effects. For a video with only monaural or stereo audio, MIMOSA automatically grounds each sound source to the corresponding sounding object in the visual scene and enables users to further validate and fix the errors in the locations of sounding objects. Users can also augment the spatial audio effect by flexibly manipulating the sounding source positions and creatively customizing the audio effect. The design of MIMOSA exemplifies a human-AI collaboration approach that, instead of utilizing state-of-art end-to-end "black-box" ML models, uses a multistep pipeline that aligns its interpretable intermediate results with the user’s workflow. A lab user study with 15 participants demonstrates MIMOSA’s usability, usefulness, expressiveness, and capability in creating immersive spatial audio effects in collaboration with users.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/cc/157782/2024

AdRecommended

Learn AI Coding at CodeNow

At a Glance

Paper Snapshot

fact_check
dataset
Source
C&C
calendar_month
Year
2024
emoji_events
Award
No award tagged
group
Authors
7 authors
sell
Subtopics
Generative AI (Text, Image, Music, Video), Music Composition & Sound Design Tools, Creative Collaboration & Feedback Systems
work
Professions
Content Creators (YouTubers, Podcasters), Musicians, DJs & Sound Designers, Film & Animation Producers
article
Content Status
Abstract only
hub
Related Papers
9 related papers