RiskRAG: A Data-Driven Solution for Improved AI Model Risk Reporting

Explainable AI (XAI)AI-Assisted Decision-Making & AutomationAI Ethics, Fairness & AccountabilitySoftware Engineers & DevelopersUI/UX DesignersAI/ML Researchers & Engineers

Research Background and Problem

  • Problem or Challenge:
    The authors point out that the risk reporting sections in current artificial intelligence (AI) model documentation are generally lacking. An analysis of 450,000 model cards revealed that only 14% included risk-related content, with 96% of those directly copying from existing documentation. This lack of diversity in content and actionable recommendations diminishes the practical value of the documentation. Furthermore, AI risk reports are often criticized for being too vague and impractical, failing to adequately meet the needs of developers and users.

  • Significance:
    Transparent and comprehensive risk reporting is critical for developing trustworthy AI systems. It not only helps developers identify potential ethical issues in models but also enables downstream applications to effectively assess and manage risks. Risk reporting also aids in meeting compliance requirements for high-risk AI systems, such as those outlined in the EU AI Act.

  • Research Motivation:
    Although previous studies have attempted to optimize model documentation formats and content generation tools, such as Risk Cards and CardGen, they often fall short of meeting the specific requirements for identifying, prioritizing, and mitigating risks associated with particular models. The authors propose the need for a data-driven and automated tool to address these shortcomings.

Solution

  • Method or Solution:
    The authors propose RiskRAG, an AI risk report generation system based on Retrieval-Augmented Generation (RAG). By combining data from model cards and real-world AI incident databases, RiskRAG automatically generates contextualized, structured risk reports tailored to the needs of various stakeholders.

  • Innovations:

    1. Combines human-written risk descriptions with retrieval-augmented generation techniques to reduce the issue of "hallucinated generation."
    2. Generates customized risk reports for specific models rather than generalized AI risks.
    3. Provides prioritized risk rankings along with actionable mitigation strategies.
    4. Integrates an AI incident database to offer real-world case references for model risk assessment.
  • Implementation Steps and Key Techniques:

    1. Retrieval Module: Compares model descriptions from existing model cards and AI incident databases to retrieve relevant risk content.
    2. Generation Module:
      • Uses generative models such as GPT-4 to extract key risks from retrieved content and categorize them into structured classifications (e.g., false positives, bias).
      • Generates practical application scenarios based on model use cases, linking each risk to specific scenarios.
      • Maps mitigation strategies to specific risks.
    3. Risk Prioritization:
      • Ranks risks based on their recurrence in different cases and the severity of actual harm caused.
    4. Output Format:
      • Provides structured risk tables (e.g., risk heatmaps) and actionable recommendations, tailored for different types of users (developers, designers, decision-makers).

Research Outcomes

  • Specific Results:
    RiskRAG surpasses traditional model card risk reports by generating more detailed and contextualized risk assessments, encouraging developers and users to make more cautious choices when selecting and deploying AI models.

  • Advantages Compared to Existing Solutions:

    1. Broader Coverage: RiskRAG captures a wider range of model-specific risks, extending beyond technical layers to include application-related risks.
    2. Actionable Strategies: The generated mitigation strategies are easy to understand and implement.
    3. Support for Prioritization: RiskRAG effectively ranks risks based on frequency and severity in real-world cases, helping users focus on critical issues.
    4. Positive User Feedback: In preliminary and final user studies, RiskRAG reports were significantly preferred by developers, designers, and media professionals over traditional model card risk sections.
  • Experimental or Evaluation Results:

    1. Risk Coverage and Depth: In user studies, RiskRAG-generated reports scored significantly higher than traditional reports in terms of coverage, specificity, and comprehensibility.
    2. Impact on Decision Quality: RiskRAG facilitated deeper risk identification and mitigation planning, promoting more cautious decision-making processes.
    3. Applicability Validation: Beyond mainstream model cards, experiments with less common models validated RiskRAG's generalizability, showing it could still provide highly relevant content.
  • Limitations and Future Directions:

    1. Model Database Limitations: Currently relies primarily on data from HuggingFace and AIID; future work should integrate more diverse risk databases (e.g., OECD AIM).
    2. Learning Curve and Familiarity: Some users require time to adapt to RiskRAG's visualized report format; a hybrid approach combining text and graphical reports could be considered.
    3. Refinement of Risk Strategies: Mitigation strategies for certain scenarios, especially for niche or emerging models, could be further improved.
    4. Interactive Design Optimization: Future work could explore dynamic, interactive risk report interfaces to enhance user experience.

Through RiskRAG, the authors provide a systematic, data-driven, and scalable solution for risk reporting in machine learning models, marking a significant advancement in AI risk management.

Quick Actions

Share

Share this page

ios_share

https://hci.top/en/papers/chi/189610/2025

AdRecommended

Learn AI Coding at CodeNow

open_in_newOpen DOI Link
DOI: https://dl.acm.org/doi/10.1145/3706598.3713979
At a Glance

Paper Snapshot

fact_check
dataset
Source
CHI
calendar_month
Year
2025
emoji_events
Award
No award tagged
group
Authors
5 authors
sell
Subtopics
Explainable AI (XAI), AI-Assisted Decision-Making & Automation, AI Ethics, Fairness & Accountability
work
Professions
Software Engineers & Developers, UI/UX Designers, AI/ML Researchers & Engineers
article
Content Status
Full text indexed
hub
Related Papers
10 related papers