Mental Models in Human-AI Interaction: Systematic Review of Empirical Methodologies and Guidelines
Authors
The notion of mental model has long been used in HCI to capture people's understanding and reasoning about computing systems. Eliciting users' mental models can explain their behaviors and attitudes toward a system—why and how they use, rely on, trust, or reject it. However, its use remains conceptually fragmented and methodologically diverse and has not been revisited in light of modern AI systems, whose opacity and newfound abilities may challenge human understanding. To address this gap, we systematically review 88 empirical studies that elicit humans’ mental models of AI systems. We extracted and analyzed how studies define and elicit mental models, the type of mental model their method presupposes, and how these vary across AI system types. Drawing from the mental model's framing in cognitive psychology and HCI, and based on descriptive and relational analysis between the variables extracted, we find that (1) mental model elicitations' goal bifurcates between system-specific evaluation and class-level probes surfacing lay theories; (2) epistemic assumptions exceed the classic functional-structural lens (how the system behaves / how it works internally) with analogical and anthropomorphic framings of AI systems; (3) elicitation methods are shaped more by system characteristics and community-specific practices than theoretical commitments, particularly for predictive and explainable AI systems and autonomous or driver-assist vehicles. We derive 9 practical guidelines to support more deliberate and reflective methods for eliciting mental models of AI systems. In doing so, we aim to reestablish continuity between the cognitive theory of mental models and their empirical use in HCI, improving the transparency and comparability of research surrounding the concept.
Research Questions / Practical Problems
Question signals indexed for this paper.
- 100%
What Did My Car Say? Impact of Autonomous Vehicle Explanation Errors and Driving Context On Comfort, Reliance, Satisfaction, and Driving Confidence
CHI '25· Automated Driving Interface & Takeover Design +2
- 83%
How Do People Rank Multiple Mutant Agents?
IUI '22· Explainable AI (XAI) +1
- 71%
People Attribute Purpose to Autonomous Vehicles When Explaining Their Behavior: Insights from Cognitive Science for Explainable AI
CHI '25· Automated Driving Interface & Takeover Design +1
- 67%
Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performance
CHI '21· Explainable AI (XAI) +1
- 67%
One AI Does Not Fit All: A Cluster Analysis of the Laypeople’s Perception of AI Roles
CHI '23· Explainable AI (XAI) +1
- 67%
"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction
CHI '23· Explainable AI (XAI) +1
- 67%
When to Explain: Modeling User Need for Explanations in Real-World Autonomous Driving
CHI '26· Explainable AI (XAI) +1
- 67%
Editable XAI: Toward Bidirectional Human-AI Alignment with Co-Editable Explanations of Interpretable Attributes
CHI '26· Explainable AI (XAI) +1
- 67%
Emergent, not Immanent: A Baradian Reading of Explainable AI
CHI '26· Explainable AI (XAI) +1
- 67%
Automated Rationale Generation: a Technique for Explainable AI and its Effects on Human Perceptions
IUI '19· Explainable AI (XAI) +1
Based on Jaccard similarity of research subtopics & professions (≥60%)