Crosscast: Adding Visuals to Audio Travel Podcasts

Museum & Cultural Heritage DigitizationInteractive Narrative & Immersive StorytellingContent Creators (YouTubers, Podcasters)Journalists & Editors

Audio travel podcasts are a valuable source of information for travelers. Yet, travel is, in many ways, a visual experience and the lack of visuals in travel podcasts can make it difficult for listeners to fully understand the places being discussed. We present Crosscast: a system for automatically adding visuals to audio travel podcasts. Given an audio travel podcast as input, Crosscast uses natural language processing and text mining to identify geographic locations and descriptive keywords within the podcast transcript. Crosscast then uses these locations and keywords to automatically select relevant photos from online repositories and synchronizes their display to align with the audio narration. In a user evaluation, we find that 85.7% of the participants preferred Crosscast generated audio-visual travel podcasts compared to audio-only travel podcasts.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/uist/42102/2020

AdRecommended

Learn AI Coding at CodeNow

open_in_newOpen DOI Link
DOI: https://dl.acm.org/doi/10.1145/3379337.3415882
At a Glance

Paper Snapshot

fact_check
dataset
Source
UIST
calendar_month
Year
2020
emoji_events
Award
No award tagged
group
Authors
3 authors
sell
Subtopics
Museum & Cultural Heritage Digitization, Interactive Narrative & Immersive Storytelling
work
Professions
Content Creators (YouTubers, Podcasters), Journalists & Editors
article
Content Status
Abstract only
hub
Related Papers
0 related papers