Towards Human-AI Deliberation: Design and Evaluation of LLM-Empowered Deliberative AI for AI-Assisted Decision-Making
Honorable MentionAuthors
Research Background and Issues
-
Identified Challenges
- Traditional AI-assisted decision-making systems provide fixed recommendations, limiting user interaction with AI to merely accepting or rejecting suggestions, especially when disagreements arise.
- Users' reliance on AI recommendations may lead to over-reliance or complete disregard (under-reliance), ultimately affecting decision quality.
- Current systems fail to adequately address and resolve partial consistency and conflicts in reasoning between humans and AI during decision-making.
-
Significance
- AI decision support has been widely applied in critical domains such as medical diagnosis, investment decision-making, and criminal justice, but deficiencies in accuracy and transparency can result in severe consequences.
- Improved human-AI collaboration can enhance overall team decision-making performance and help address ethical and fairness issues in real-world applications.
-
Research Motivation and Related Work
- This study draws inspiration from human deliberation theories and the "Weight of Evidence (WoE)" framework, aiming to introduce deeper interactions between AI and users to improve critical thinking and trust.
- Current explainable AI (XAI) primarily simplifies AI decision processes by presenting model explanations, but research on enhancing AI-assisted decision-making through dynamic discussion and conflict resolution remains limited.
Solution
-
Proposed Approach This paper introduces the "Human-AI Deliberation" framework and develops its core component, "Deliberative AI," which leverages large language models (LLMs) for dynamic interaction to address inconsistencies between human and AI perspectives.
-
Innovations
- Support for Dimension-Level Opinion Expression and Updates: Human and AI can exchange and adjust opinions on specific dimensions (e.g., preference weights) rather than merely accepting or rejecting AI's overall suggestions.
- Encouraging Dynamic and Structured Discussions: A dialogue interface is designed to facilitate human-AI collaboration, supporting mutual questioning, evidence verification, and opinion updates.
- Integration of LLMs and Domain-Specific Models (DS Models): LLMs enable natural language interaction, while domain-specific models ensure reliable and accurate information provision.
-
Implementation Steps and Techniques
- Expression and Alignment of Ideas: Domain models use SHAP explanations to generate dimension-specific feature weights (Weight of Evidence, WoE). Users are also required to clearly express their opinions on each dimension.
- Human-AI Discussion: The system uses LLMs to identify human intentions and dimension-specific queries, guiding the dialogue and automatically invoking domain models to extract specific evidence (e.g., data distributions, global correlations).
- Opinion Updates: Formulas quantify the strength of users' arguments and AI uncertainty, dynamically adjusting AI perspectives for more proactive responses.
- Interface Design: An intuitive interactive interface is developed to support dimension-level detailed discussions, overall decision updates, and summary viewing.
Research Outcomes
-
Specific Results
- Through application in the research task "graduate admission evaluation," Human-AI Deliberation significantly improved collaborative decision accuracy.
- Compared to traditional Explainable AI (XAI), this approach effectively reduced over-reliance on erroneous AI recommendations while minimizing ineffective behaviors caused by unreliable AI explanations.
-
Experimental Performance and Advantages
- Experiments show that compared to traditional XAI, "Deliberative AI" improved participants' decision accuracy (accuracy increased from 52.4% to 59.8%).
- Reduced over-reliance: Over-reliance rates dropped from 65% in traditional systems to 47% with this method.
- Task complexity did not significantly increase, and users positively acknowledged the deep interaction and collaboration process with AI.
-
Limitations and Future Directions
- Context and Task Applicability: The research task (graduate admission case) is based on synthetic datasets, which may differ from real-world high-stakes decision scenarios.
- User Experience and Satisfaction: Some user feedback indicated that the discussion process was lengthy, increasing decision-making burdens, necessitating optimization of input modules and dialogue generation speed.
- Applicability to Multi-Modal Data Tasks: The current study primarily focuses on tabular data; future research should explore extensions to tasks involving image or text analysis, such as through vision-language models.
- Ethics and Transparency: Users expressed concerns about the transparency of AI's opinion update mechanisms, requiring clearer presentation of the basis for AI adjustments in future work.
Conclusion
The "Human-AI Deliberation" framework proposed in this paper offers a novel perspective for AI-assisted decision-making by introducing dynamic, detailed user participation and AI interaction into critical domains. Experiments demonstrate its potential to significantly improve decision quality and provide initial insights into user perceptions of Deliberative AI. Future research should further expand task diversity and optimize human-AI interaction efficiency to adapt to broader real-world application scenarios.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- Is there room to improve traditional AI decision support systems when human and AI views disagree?Category: AI Decision Support and Reliance BehaviorSimilar questionsarrow_forward
- How can dynamic discussion and conflict resolution enhance human-AI decision collaboration?Category: AI Decision Support and Reliance BehaviorSimilar questionsarrow_forward
- Can human-AI collaboration frameworks incorporating large language models reduce users' over-reliance on incorrect AI advice?Category: AI Decision Support and Reliance BehaviorSimilar questionsarrow_forward
Practical Problems
1- AI decision systems cannot resolve human-AI disagreement, affecting decision quality.Category: AI Decision Support and Reliance BehaviorSimilar questionsarrow_forward
- 100%
Effects of Communication Directionality and AI Agent Differences in Human-AI Interaction
CHI '21· Human-LLM Collaboration +1
- 100%
AI Knowledge: Improving AI Delegation through Human Enablement
CHI '23· Human-LLM Collaboration +1
- 100%
A Survey on Interactive Reinforcement Learning: Design Principles and Open Challenges
DIS '20· Human-LLM Collaboration +1
- 80%
You Complete Me: Human-AI Teams and Complementary Expertise
CHI '22· Human-LLM Collaboration +1
- 80%
Competent but Rigid: Identifying the Gap in Empowering AI to Participate Equally in Group Decision-Making
CHI '23· Human-LLM Collaboration +1
- 80%
Why Johnny Can’t Prompt: How Non-AI Experts Try (and Fail) to Design LLM Prompts
CHI '23· Human-LLM Collaboration +1
- 80%
Automatic Macro Mining from Interaction Traces at Scale
CHI '24· Human-LLM Collaboration +1
- 80%
From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered Analysis
CHI '25· Human-LLM Collaboration +2
- 80%
Which Contributions Deserve Credit? Perceptions of Attribution in Human-AI Co-Creation
CHI '25· Human-LLM Collaboration +2
- 80%
Satisficing vs. Maximizing in Prompt Writing: Trait and Task Effects in Human–AI Interaction
CHI '26· Generative AI (Text, Image, Music, Video) +2
Based on Jaccard similarity of research subtopics & professions (≥60%)