Explanations, Fairness, and Appropriate Reliance in Human-AI Decision-Making
Honorable MentionAuthors
Title of the Paper
Explanations, Fairness, and Appropriate Reliance in Human-AI Decision-Making
Paper Information
- Research Domain: Human-Computer Interaction and Fairness in Artificial Intelligence
- Keywords: Human-AI Interaction, AI-Assisted Decision-Making, Appropriate Reliance, Explainable AI, Algorithmic Fairness, Fairness Perception
Research Background and Problem Statement
-
Problems and Challenges:
- With the widespread application of AI systems in critical domains, algorithmic biases in their decision-making recommendations may lead to unfair outcomes.
- Many studies suggest that explainability can help human decision-makers identify biases, but there is insufficient empirical evidence to confirm whether existing explanation techniques truly possess this capability.
-
Significance:
- Fairness and accuracy are core issues in AI-assisted decision-making, particularly in high-stakes domains such as finance and recruitment, where algorithmic biases may negatively impact specific groups, such as gender discrimination in occupational predictions.
-
Motivation and Related Work:
- Existing research primarily focuses on the impact of explanations on fairness perception and human trust in AI, but the specific influence of explanations on distributive fairness has not been thoroughly explored.
- The authors further analyze how explanations affect humans' ability to revise AI recommendations and how such behavior improves or worsens distributive fairness.
Proposed Solution
-
Proposed Approach:
- The authors designed a randomized online experiment to investigate how feature-based explanations influence humans' ability to enhance distributive fairness, while also exploring the behavioral mechanisms (e.g., fairness perception and reliance behavior) that affect this outcome.
-
Innovations:
- This is the first systematic analysis of the impact of explanations on distributive fairness and their relationship with fairness perception and reliance behavior, addressing a research gap in the field.
- The study operationalizes this analysis by constructing two AI models (one based on task-relevant words and the other on gender-related words) and providing explanation-based recommendations.
-
Implementation Steps and Key Techniques:
- Experimental Design: In an occupational prediction scenario, participants were asked to judge whether biographical data belonged to a professor or a teacher based on AI predictions and explanations. The models involved gender-related and task-related features.
- Dataset Selection: The BIOS public dataset was used, containing biographical texts along with corresponding occupation and gender information.
- Evaluation Metrics:
- Measure types of human reliance behavior on AI recommendations (e.g., corrective reliance, detrimental reliance).
- Analyze error rate disparities based on gender (e.g., "teacher -> professor" error differences).
- Collect fairness perception data, quantified using Likert scale survey responses.
Research Findings
-
Key Findings:
- Explanations Do Not Improve Accuracy: There was no significant difference in participants' decision accuracy with or without explanations.
- Impact on Reliance Behavior:
- Gender-related explanations led to more frequent rejection of AI recommendations, but these rejections were not related to the correctness of the recommendations.
- Task-related feature explanations encouraged participants to follow AI recommendations but could reinforce gender stereotypes.
- Impact on Fairness:
- Gender-related feature explanations reduced gender error rate disparities (improving distributive fairness), while task-related feature explanations increased disparities (worsening fairness).
- These changes were primarily due to shifts in error types rather than an enhanced ability to correct erroneous recommendations.
-
Comparison with Existing Solutions:
- The study highlights that existing feature-based explanation techniques may not be suitable for improving distributive fairness.
- It emphasizes the complex relationship between fairness perception and improvements in distributive fairness, revealing that relying solely on fairness perception is insufficient to measure behavioral outcomes.
-
Limitations and Future Directions:
- The experiment did not collect fairness perception data at the instance level; future research could investigate how instance-level perceptions influence overall perceptions and behaviors.
- Further research is needed on individual differences in behavioral responses to explanations, such as tailoring explanation strategies based on participants' gender, cultural background, etc.
- Explore new methods to directly convey a model's fairness information to humans, rather than relying solely on feature-based explanations.
Conclusion and Recommendations
-
Significance:
- The authors propose a novel evaluation pathway to study the impact of explanations on fairness and behavior, exploring potential issues from multiple dimensions.
- They emphasize the importance of designing explanations with specific goals in mind and call for broader consideration of how to provide effective transparency information.
-
Practical Recommendations:
- When designing AI explanation mechanisms, their actual impact on reliance behavior and fairness metrics should be evaluated, rather than being limited to perception data.
- Shift from current feature-based explanation methods to approaches that directly convey system-wide fairness information, aiding practical decision-making and the design of socio-technical systems.
This study provides valuable insights into the exploration of fairness issues and potential solutions in human-AI collaboration.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How do feature explanations affect human reliance on AI recommendations and the ability to correct them?Category: Fairness Perception, Resource Allocation, and Interaction PresentationSimilar questionsarrow_forward
- Do feature explanations help improve allocative fairness (fairness of allocation outcomes)?Category: Fairness Perception, Resource Allocation, and Interaction PresentationSimilar questionsarrow_forward
- What is the relationship between perceived fairness and improvements in allocative fairness?Category: Fairness Perception, Resource Allocation, and Interaction PresentationSimilar questionsarrow_forward
Practical Problems
1- AI systems may amplify bias in tasks, leading to unfair allocations, such as gender discrimination in occupational prediction.Category: Fairness Perception, Resource Allocation, and Interaction PresentationSimilar questionsarrow_forward
- 86%
Knowing About Knowing: An Illusion of Human Competence Can Hinder Appropriate Reliance on AI Systems
CHI '23· Explainable AI (XAI) +2
- 75%
What Lies Beneath? Exploring the Impact of Underlying AI Model Updates in AI-Infused Systems
CHI '25· Generative AI (Text, Image, Music, Video) +2
- 75%
Emulating Aggregate Human Choice Behavior and Biases with GPT Conversational Agents
CHI '26· Human-LLM Collaboration +3
- 71%
Farsight: Fostering Responsible AI Awareness During AI Application Prototyping
CHI '24· Explainable AI (XAI) +2
- 71%
Trust in AI-assisted Decision Making: Perspectives from Those Behind the System and Those for Whom the Decision is Made
CHI '24· Explainable AI (XAI) +2
- 71%
The Who in XAI: How AI Background Shapes Perceptions of AI Explanations
CHI '24· Explainable AI (XAI) +1
- 71%
User Characteristics in Explainable AI: The Rabbit Hole of Personalization?
CHI '24· Explainable AI (XAI) +1
- 71%
Trusting Autonomous Teammates in Human-AI Teams - A Literature Review
CHI '25· Explainable AI (XAI) +2
- 71%
I Can Do Better Than Your AI: Expertise and Explanations
IUI '19· Explainable AI (XAI) +1
- 67%
Agent-Supported Foresight for AI Systemic Risks: AI Agents for Breadth, Experts for Judgment
CHI '26· Explainable AI (XAI) +4
Based on Jaccard similarity of research subtopics & professions (≥60%)