Rethinking Human-AI Collaboration in Complex Medical Decision Making: A Case Study in Sepsis Diagnosis
Authors
Xuhai "Orson" Xu
Massachusetts Institute of TechnologyTitle of the Paper
Rethinking Human-AI Collaboration in Complex Medical Decision-Making: A Case Study on Sepsis Diagnosis
Paper Information
- Subject Area: Human-AI Collaboration, Medical Artificial Intelligence
- Keywords: Human-AI Collaboration, Medical Decision Support, Sepsis Diagnosis, Explainable AI (XAI), Clinical Applications, Uncertainty Visualization, Experimental Systems, Time-Sensitive Decision-Making, High-Risk Decisions
Research Background and Problem
-
Key Issues:
- Many current AI models perform well on academic datasets but are often abandoned and difficult to deploy in human experts' workflows.
- Existing AI systems (e.g., the Epic Sepsis Module, ESM) focus on supporting the final decision-making stage (e.g., risk scoring) rather than addressing the more critical intermediate stages of clinical workflows, failing to meet clinical needs effectively.
- Clinical scenarios (e.g., sepsis diagnosis) are characterized by high risk, high uncertainty, and time sensitivity. AI systems provide insufficient support for these challenges and are often perceived as "challengers" rather than "collaborators."
-
Research Motivation:
- Sepsis is a high-risk and rapidly progressing condition where early diagnosis is critical for patient survival.
- Human experts require more intuitive and actionable AI support.
- User feedback on existing AI systems highlights the need for redesign to better integrate into real-world workflows.
-
Related Work:
- Current XAI (Explainable AI) research primarily focuses on improving algorithm transparency and interpretability but has not adequately addressed clinicians' needs for AI interaction during early diagnosis.
- The design of AI in other medical decision-support systems lacks attention to clinical uncertainty and human expert requirements.
Solution
-
Method Overview:
- Propose a human-centered AI system named "SepsisLab," which reimagines AI's role from supporting "final decisions" to assisting in "intermediate decision stages."
- SepsisLab optimizes human-AI collaboration through five design strategies, including future prediction, uncertainty visualization, and experimental recommendations.
-
Design Highlights:
- Providing Future Trend Predictions: Predict not only current risks but also changes in risk levels and uncertainty ranges over the next few hours.
- Generating Experimental Recommendations: Suggest blood tests with the highest information gain to clinicians to reduce diagnostic uncertainty.
- Visualizing Uncertainty: Use graphical representations to display the temporal changes in risk prediction uncertainty, making predictions more intuitive and credible.
- Building Interactive Counterfactual Simulations: Allow clinicians to adjust hypothetical scenarios to explore how changes in variables might impact risk scores.
- Repositioning the Role of AI: Shift AI from being a "decision authority" to a "collaborator" that supports clinicians' reasoning processes and information analysis.
-
Implementation Steps and Techniques:
- Utilize LSTM deep learning models for time-series predictions, generating future forecasts based on patients' historical data.
- Employ Monte Carlo Simulation (MCS) to estimate uncertainty in risk predictions and support experimental recommendations.
- Develop an interactive front-end user interface to enable visualization and counterfactual operations (implemented using the React framework).
Research Outcomes
-
Specific Results:
- SepsisLab System Prototype:
- Implemented dynamic prediction of sepsis diagnosis, uncertainty range visualization, and experimental recommendation functionalities.
- Provided intuitive support for clinicians across multiple stages, including "hypothesis generation - data collection - hypothesis testing."
- Enhanced Collaboration and Transparency Experience:
- User evaluations indicated that the new design significantly improved human-AI team collaboration and reduced the perception of AI as a "decision authority" challenging clinicians.
- System Algorithm Performance Evaluation:
- Experiments on the MIMIC-III dataset showed that with a 9.6% increase in necessary experimental values, the recommendation method achieved prediction performance close to that of the complete dataset.
- SepsisLab System Prototype:
-
Advantages Analysis:
- Comparison with Existing ESM Module: Avoided the ESM module's overemphasis on a single final score, reducing false alarms and ambiguous feedback.
- Positioned AI as a "collaborator," mitigating the risk of clinicians distrusting or abandoning AI tools.
-
Limitations and Future Directions:
- Limited Scope of User Studies: The system was evaluated by only six clinicians; future work requires larger-scale and more diverse clinical trials.
- Challenges in Real-World Integration: SepsisLab has not yet been integrated into the actual workflows of U.S. Epic hospitals, requiring extended evaluation periods.
- Improvement Directions:
- Enhance the system's interactive interface to prevent data or information overload.
- Expand to other complex medical scenarios and non-medical high-risk decision-making domains, such as emergency response and military planning.
- Further optimize algorithm performance and develop more transparent model interpretation methods.
Conclusion
SepsisLab redefines the collaboration model of medical AI in complex scenarios, shifting its role from a "decision provider" to a "decision supporter." Its design principles are not only applicable to sepsis diagnosis but also provide a reference for building more effective human-AI collaboration models in other domains.
Research Questions / Practical Problems
Question signals indexed for this paper.
Research Questions
3- How can AI systems be improved to support intermediate decision stages in medical diagnosis rather than focusing only on final decisions?Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
- How can human-centered AI tools be designed to enhance physician trust and collaboration in high-risk, high-uncertainty, and time-sensitive clinical scenarios?Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
- What visualization and interaction strategies can help physicians more intuitively understand AI predictions and uncertainty?Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
Practical Problems
1- Existing AI tools provide limited support for physicians making rapid sepsis diagnoses.Category: Uncertainty Communication and Calibrated RelianceSimilar questionsarrow_forward
- 86%
Augmenting Pathologists with NaviPath: Design and Evaluation of a Human-AI Collaborative Navigation System
CHI '23· Explainable AI (XAI) +2
- 75%
Rapid Assisted Visual Search: Supporting Digital Pathologists with Imperfect AI
IUI '21· EV Charging & Eco-Driving Interfaces +3
- 71%
Ambiguity-aware AI Assistants for Medical Data Analysis
CHI '20· Explainable AI (XAI) +1
- 71%
EXMOS: Explanatory Model Steering through Multifaceted Explanations and Data Configurations
CHI '24· Explainable AI (XAI) +1
- 71%
Visual-Conversational Interface for Evidence-Based Explanation of Diabetes Risk Prediction
CUI '25· Explainable AI (XAI) +2
- 63%
Accurate Insights, Trustworthy Interactions: Designing a Collaborative AI-Human Multi-Agent System with Knowledge Graph for Diagnosis Prediction
CHI '25· Brain-Computer Interface (BCI) & Neurofeedback +2
- 63%
VMS: Interactive Visualization to Support the Sensemaking and Selection of Predictive Models
IUI '24· Explainable AI (XAI) +2
Based on Jaccard similarity of research subtopics & professions (≥60%)