Perceptions of the Fairness Impacts of Multiplicity in Machine Learning
Authors
Research Background and Problem
- Identified Issues or Challenges: The authors explore the impact of "diversity" on the fairness of machine learning (ML), where diversity refers to the existence of multiple different models with comparable predictive accuracy. This diversity can lead to the same input data being predicted differently by different models, potentially creating risks to procedural fairness. However, current literature has not investigated whether "stakeholders" also perceive this risk.
- Significance: When ML is applied to high-stakes decisions (e.g., social services, criminal justice, and healthcare), such arbitrariness can have substantial impacts on individuals, such as losing job opportunities or access to healthcare. This directly affects public trust and acceptance of ML systems.
- Research Motivation and Related Work: Existing studies from philosophical and computer science perspectives suggest that diversity may pose risks to fairness. However, there is a lack of direct dialogue with ordinary stakeholders (e.g., decision subjects). Given that the outputs of ML systems affect public interests, understanding their perceptions is crucial.
Solution
- Method or Solution: Through a survey study, the authors measured the perceptions of the general public (i.e., decision recipients) regarding the impact of diversity on fairness in ML systems and their preferences for addressing diversity issues.
- Innovative Contribution: This is the first systematic investigation into non-technical stakeholders' views on the fairness implications of diversity, bridging the gap between theoretical assumptions and public perceptions.
- Implementation Steps:
- Conduct a preliminary study to select several tasks.
- Use a main study to investigate how diversity affects participants' perceptions of ML fairness.
- Collect participants' preferences and compare six diversity resolution methods (e.g., randomization, human expert intervention).
- Analyze response patterns using qualitative and quantitative methods, and examine how task characteristics (e.g., risk and reward frameworks) moderate participants' responses.
Research Findings
- Specific Findings:
- Ordinary stakeholders generally do not perceive diversity as a significant threat to the fairness of ML systems, but they exhibit clear preferences for resolving diversity issues.
- The most preferred approach is human expert intervention, while the least preferred is randomization.
- Preferences vary depending on task characteristics; for instance, in high-risk tasks, participants favor human decision-making.
- Comparison with Existing Solutions:
- Current ML practices often directly select a single well-performing model while ignoring diversity, but this approach (ignore) is perceived as unfair by the public.
- Philosophers' proposal of randomization as a fairness solution is also not accepted by the public.
- Experimental or Evaluation Results:
- Preference for "human decision-making" (marginal mean of 0.795, significantly higher than the random expectation value of 0.5).
- In high-risk tasks, there is a stronger preference for "human" or "complex models" (e.g., ensemble methods), while randomization is more tolerable in low-risk tasks.
- Limitations and Future Directions:
- Participants' understanding of diversity could be further improved through education, as some results may be constrained by their technical knowledge.
- The current study is based on hypothetical tasks; future research could focus on real-world scenarios to collect data with higher ecological validity.
- The reasons behind participants' preferences, particularly their strong aversion to randomization (which contradicts philosophical recommendations), warrant further exploration.
Conclusion
This study challenges the assumptions of current ML practices and philosophical recommendations, particularly in addressing the issue of predictive diversity. While the presence of diversity does not significantly reduce participants' perceptions of fairness, they exhibit clear preferences for resolving diversity issues. This suggests that ML developers need to address diversity issues more transparently and purposefully, while considering stakeholder involvement and perceptions in system design. The research provides a new perspective on the societal dimensions of ML fairness and lays a foundation for enhancing public trust and acceptance of algorithmic decision-making in the future.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- Do ordinary stakeholders (e.g., those affected by decisions) perceive prediction diversity as a fairness risk in machine learning systems?Category: Fairness, Bias, and RepresentationSimilar questionsarrow_forward
- Which solutions do laypeople prefer when addressing prediction diversity in machine learning?Category: Fairness, Bias, and RepresentationSimilar questionsarrow_forward
- How do task characteristics (e.g., risk-reward framing) affect public preferences on prediction diversity issues?Category: Fairness, Bias, and RepresentationSimilar questionsarrow_forward
Practical Problems
1- Lay users distrust machine learning predictions and worry about fairness issues.Category: Fairness, Bias, and RepresentationSimilar questionsarrow_forward
- 100%
"Everyone wants to do the model work, not the data work": Data Cascades in High-Stakes AI
CHI '21· Explainable AI (XAI) +1
- 100%
Fairness Evaluation in Text Classification: Machine Learning Practitioner Perspectives of Individual and Group Fairness
CHI '23· Explainable AI (XAI) +1
- 80%
Beyond Expertise and Roles: A Framework to Characterize the Stakeholders of Interpretable Machine Learning and their Needs
CHI '21· Explainable AI (XAI) +2
- 80%
User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
CHI '25· Explainable AI (XAI) +2
- 80%
More than Marketing? On the Information Value of AI Benchmarks for Practitioners
IUI '25· Explainable AI (XAI) +1
- 75%
Designing Interactive Explainable AI Tools for Algorithmic Literacy and Transparency
DIS '24· Explainable AI (XAI) +1
- 75%
Conversational Explanations: Discussing Explainable AI with Non-AI Experts
IUI '25· Explainable AI (XAI)
- 67%
Do Expressions Change Decisions? Exploring the Impact of AI's Explanation Tone on Decision-Making
CHI '25· Explainable AI (XAI) +2
- 67%
Explaining Models: An Empirical Study of How Explanations Impact Fairness Judgment
IUI '19· Explainable AI (XAI) +2
- 67%
The Impact of Explanations on Fairness in Human-AI Decision-Making: Protected vs Proxy Features
IUI '24· Explainable AI (XAI) +2
Based on Jaccard similarity of research subtopics & professions (≥60%)