Empowering Medical Data Labeling for Non-Experts with DANNY: Enhancing Accuracy and Mitigating Over-Reliance on AI
Authors
Economic constraints on recruiting experts hinder efforts to build qualified datasets for utilizing AI in professional domains (e.g., medical diagnosis), which could provide societal benefits. To solve this issue, previous studies introduced crowdsourcing and AI to enable non-experts to perform expert-level data labeling. Yet, they encountered three challenges: 1) the limited applicability of crowdsourcing in less specialized domains (e.g., identifying animal species); 2) the chicken-and-egg problem, a paradox where high-performance AI is required to build a dataset to train such AI; and 3) over-reliance on AI, where non-experts, lacking expertise, may incorrectly label data when guided by sub-optimal AI. To address this, we introduce DANNY (Data ANnotation for Non-experts made easY), an AI-based tool designed to help non-experts label an arthritis dataset, aiming to increase labeling accuracy and mitigate over-reliance on AI. By externalizing a cognitive forcing intervention to foster critical thinking, DANNY provides two visualizations: 1) the Criteria phase, where non-experts define criteria across four arthritis features, and 2) the Correction phase, where they refine these criteria by comparing them to AI suggestions. In a study with 28 participants, DANNY users achieved higher accuracy and a more appropriate reliance on AI dependency than control groups. A follow-up study with 12 participants demonstrates how DANNY can be used to improve AI with an ensemble method. Our findings contribute new insights into using AI to support non-experts in labeling domain-specific data when expert resources are limited.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can non-professionals avoid over-reliance on AI when annotating highly specialized domain data?Category: Uncertainty in Active Learning and Machine TeachingSimilar questionsarrow_forward
- In medical image annotation tasks, can cognitive intervention methods based on dual-process theory improve annotation quality?Category: Uncertainty in Active Learning and Machine TeachingSimilar questionsarrow_forward
- How can limited-capability AI be effectively leveraged to help non-experts complete dataset annotation while optimizing annotation efficiency and quality?Category: Uncertainty in Active Learning and Machine TeachingSimilar questionsarrow_forward
Practical Problems
1- Non-professionals struggle to annotate domain-specific datasets at high quality.Category: Uncertainty in Active Learning and Machine TeachingSimilar questionsarrow_forward
No related papers with ≥60% similarity
Based on Jaccard similarity of research subtopics & professions (≥60%)