Investigating the Effect of the Multiple Comparisons Problem in Visual Analysis

Uncertainty VisualizationVisualization Perception & CognitionUniversity Professors & ResearchersData Scientists & AnalystsStatisticians & Data Scientists

The goal of a visualization system is to facilitate dataset-driven insight discovery. But what if the insights are spurious? Features or patterns in visualizations can be perceived as relevant insights, even though they may arise from noise. We often compare visualizations to a mental image of what we are interested in: a particular trend, distribution or an unusual pattern. As more visualizations are examined and more comparisons are made, the probability of discovering spurious insights increases. This problem is well-known in Statistics as the multiple comparisons problem (MCP) but overlooked in visual analysis. We present a way to evaluate MCP in visualization tools by measuring the accuracy of user reported insights on synthetic datasets with known ground truth labels. In our experiment, over 60% of user insights were false. We show how a confirmatory analysis approach that accounts for all visual comparisons, insights and non-insights, can achieve similar results as one that requires a validation dataset.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/chi/5772/2018

AdRecommended

Learn AI Coding at CodeNow

At a Glance

Paper Snapshot

fact_check
dataset
Source
CHI
calendar_month
Year
2018
emoji_events
Award
No award tagged
group
Authors
4 authors
sell
Subtopics
Uncertainty Visualization, Visualization Perception & Cognition
work
Professions
University Professors & Researchers, Data Scientists & Analysts, Statisticians & Data Scientists
article
Content Status
Abstract only
hub
Related Papers
0 related papers