Leveraging the Twitch Platform and Gamification to Generate Home Audio Datasets

Live Streaming & Spectating ExperienceOpen-Source Collaboration & Code ReviewCitizen Science & Crowdsourced DataEsports Players & Live StreamersContent Creators (YouTubers, Podcasters)Amazon Mechanical Turk Workers

Training AI systems requires large datasets. While there are a range of existing methods for collecting such data, such as paid work on crowdsourcing platforms, the strengths and weaknesses of each method leads us to believe that new, complementary methods are needed. The Polyphonic project contributes a novel method for collecting real-world data by piggybacking on game streaming communities such as Twitch, which capture over a trillion minutes of viewer attention a year. By embedding activities within the sociotechnical context of the stream, we can leverage some of this attention for data collection and processing. In this paper, we describe the design and implementation of a proof-of-concept system for collecting home audio data. We conducted a field study in four live streams and found that our proof-of-concept effectively supports data capture. We also contribute further design insights about stream-based data collection systems.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/dis/60123/2021

AdRecommended

Learn AI Coding at CodeNow

open_in_newOpen DOI Link
DOI: https://dl.acm.org/doi/10.1145/3461778.3462097
At a Glance

Paper Snapshot

fact_check
dataset
Source
DIS
calendar_month
Year
2021
emoji_events
Award
No award tagged
group
Authors
4 authors
sell
Subtopics
Live Streaming & Spectating Experience, Open-Source Collaboration & Code Review, Citizen Science & Crowdsourced Data
work
Professions
Esports Players & Live Streamers, Content Creators (YouTubers, Podcasters), Amazon Mechanical Turk Workers
article
Content Status
Abstract only
hub
Related Papers
0 related papers