Crowdsourcing Multi-label Audio Annotation Tasks with Citizen Scientists

Data StorytellingCrowdsourcing Task Design & Quality ControlCitizen Science & Crowdsourced Data

Annotating rich audio data is an essential aspect of training and evaluating machine listening systems. We approach this task in the context of temporally-complex urban soundscapes, which require multiple labels to identify overlapping sound sources. Typically this work is crowdsourced, and previous studies have shown that workers can quickly label audio with binary annotation for single classes. However, this approach can be difficult to scale when multiple passes with different focus classes are required to annotate data with multiple labels. In citizen science, where tasks are often image-based, annotation efforts typically label multiple classes simultaneously in a single pass. This paper describes our data collection on the Zooniverse citizen science platform, comparing the efficiencies of different audio annotation strategies. We compared multiple-pass binary annotation, single-pass multi-label annotation, and a hybrid approach: hierarchical multi-pass multi-label annotation. We discuss our findings, which support using multi-label annotation, with reference to volunteer citizen scientists' motivations.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/chi/7922/2019

AdRecommended

Learn AI Coding at CodeNow

At a Glance

Paper Snapshot

fact_check
dataset
Source
CHI
calendar_month
Year
2019
emoji_events
Award
No award tagged
group
Authors
5 authors
sell
Subtopics
Data Storytelling, Crowdsourcing Task Design & Quality Control, Citizen Science & Crowdsourced Data
work
Professions
—
article
Content Status
Abstract only
hub
Related Papers
0 related papers