Comparing Apples and Oranges: Human and Computer Clustered Affinity Diagrams Under the Microscope
Authors
Document Title
Comparing Apples and Oranges: Human and Computer Clustered Affinity Diagrams Under the Microscope
Document Information
- Domain: Human-Computer Interaction and User Research
- Keywords: Text mining, affinity diagrams, automation, design research, clustering, user-centered design, fastText, human-computer collaboration, natural language processing, design efficiency
Research Background and Problem
-
Problems and Challenges:
- Affinity diagrams are essential tools in user-centered design for analyzing and organizing user statements or observations. However, manually creating affinity diagrams is time-consuming and often subjective due to differences among design team members.
- With advancements in text mining and neural networks for processing qualitative data, researchers aim to explore how algorithms can support the creation of affinity diagrams. However, effectively evaluating the quality of algorithm-generated affinity diagrams remains a challenge.
-
Research Importance:
- Affinity diagrams help design teams extract key insights from user data, supporting the generation of design inspiration. Discovering efficient and accurate methods for creating affinity diagrams would bring significant value to practice.
- The rapid development of AI technologies offers potential technical support for optimizing affinity diagram generation, which is crucial for reducing time costs and improving user research and design efficiency.
-
Research Motivation:
- To explore how automation technologies, particularly text mining models, can enhance the user research process.
- To evaluate the effectiveness of algorithm-generated affinity diagrams through quantitative and qualitative methods, identify limitations, and propose directions for improvement.
Solution
-
Methods or Solutions:
- This study compares seven text mining models (e.g., fastText, word2vec, LDA) and selects fastText as the optimal model for generating affinity diagrams.
- In the experiment, a comparative study was designed using pre-clustered and randomly ordered user data to generate affinity diagrams, exploring the support provided by fastText-generated clusters to design teams.
- Multiple metrics (technical, psychological, and performance-related) were used to quantify algorithm performance, combined with qualitative feedback from design teams and experts to analyze its impact on the design process.
-
Innovations:
- Systematic comparison of language model-based clustering (e.g., fastText) with human-generated affinity diagrams, and the proposal of evaluation metrics tailored for affinity diagrams.
- Insights into why current automated affinity diagrams fail to effectively assist design teams, providing a basis for improving algorithms and design support tools.
-
Implementation Steps and Key Technologies:
- Model Comparison: Evaluate the performance of seven different text mining models, including traditional frequency vector models and context prediction models.
- Experiment Design: Invite design teams to construct affinity diagrams using randomly ordered and fastText pre-clustered data, studying differences in efficiency and quality.
- Expert Evaluation: Experienced design experts provide subjective quality ratings for both automated and human-generated affinity diagrams.
- Qualitative Analysis: Collect feedback from design teams and experts on the use of automated affinity diagrams, analyzing algorithm effectiveness and associated issues.
Research Results
-
Specific Results:
- Among the seven text mining models evaluated, fastText performed best in semantic similarity tasks (Spearman’s ρ = 0.4431) and was selected as the primary tool for generating affinity diagrams.
- The average overlap index between automated affinity diagrams generated by fastText and human-generated diagrams was 0.694 (SD = 0.034), but the Jaccard index between clusters was relatively low (M = 0.30), indicating significant differences.
- In student team experiments, pre-clustered data did not improve efficiency; instead, it led to more discussions about technical feasibility and skepticism toward the algorithm.
-
Advantages and Limitations:
- Advantages: Technically, fastText can process large amounts of data quickly, and its subword handling capability performs well in cases of high linguistic complexity.
- Limitations:
- FastText clustering is primarily based on keywords, lacking deep semantic understanding, which prevents the generated affinity diagrams from fully meeting designers' needs for key insights.
- Design teams' lack of understanding of automated technology and distrust in clustering results limits practical application.
-
Experiment and Evaluation Results:
- Design teams reported low satisfaction with pre-clustered data, showing no significant improvement in subjective workload or task completion progress.
- Experts rated fastText-generated affinity diagrams poorly, citing a lack of intrinsic semantic consistency and inability to directly support design insights.
- Qualitative interviews revealed that both design teams and experts believe current technical support requires greater transparency and customization improvements.
-
Future Directions:
- Develop semi-automated support tools that integrate designer interaction rather than fully replacing manual operations with automation.
- Achieve higher levels of semantic understanding (e.g., capturing user statements' emotions and context) to enhance clustering depth and consistency.
- Further explore algorithm parameter optimization and context-based clustering recommendation systems tailored to different design scenarios.
Conclusion: This study finds that existing text mining technologies have limited practical effectiveness in generating affinity diagrams, but their potential can be realized through targeted follow-up research and tool design improvements.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- Which text mining models perform best for generating affinity diagrams?Category: Affinity Diagram Generation and Design Team AnalysisSimilar questionsarrow_forward
- How do algorithm-generated and human-generated affinity diagrams differ in semantic consistency and design support?Category: Affinity Diagram Generation and Design Team AnalysisSimilar questionsarrow_forward
- How do design teams perceive pre-clustered data generated by fastText in terms of task efficiency and design inspiration?Category: Affinity Diagram Generation and Design Team AnalysisSimilar questionsarrow_forward
Practical Problems
1- Design teams find manual affinity diagram creation time-consuming and subjectively inconsistent.Category: Affinity Diagram Generation and Design Team AnalysisSimilar questionsarrow_forward
- 75%
Storyboard-Based Empirical Modeling of Touch Interface Performance
CHI '18· Prototyping & User Testing
- 75%
Steering through Successive Objects
CHI '18· Prototyping & User Testing
- 75%
Applied Sketching in HCI: Hands-on Course of Sketching Techniques
CHI '18· Prototyping & User Testing
- 75%
GUIComp: A GUI Design Assistant with Real-Time, Multi-Faceted Feedback
CHI '20· Prototyping & User Testing
- 75%
Interaction Substrates: Combining Power and Simplicity in Interactive Systems
CHI '25· Prototyping & User Testing
- 75%
Interactive Layout Transfer
IUI '21· Prototyping & User Testing
- 67%
Screen2Vec: Semantic Embedding of GUI Screens and GUI Components
CHI '21· Explainable AI (XAI) +2
- 67%
SlideAudit: A Dataset and Taxonomy for Automated Evaluation of Presentation Slides
UIST '25· Explainable AI (XAI) +2
- 60%
Balanced Interaction Design
CHI '18· Participatory Design +1
- 60%
The Algorithm and the User: How Can HCI Use Lay Understandings of Algorithmic Systems?
CHI '18· Explainable AI (XAI) +1
Based on Jaccard similarity of research subtopics & professions (≥60%)