MIT AI Risk Repository · domain 7: AI system safety, failures, & limitations

7.3 Lack of capability or robustness

AI systems that fail to perform reliably or effectively under varying conditions, exposing them to errors and failures that can have significant consequences, especially in critical applications or areas that require moral reasoning.

Risk entries
126
Frameworks citing it
12
Recorded incidents
305
Incidents since 2020
207
Causal entity (risk entries)
Causal entity (risk entries) 81 0 AI: 81 AI 81 Human: 22 Human 22 Other: 20 Other 20 Not coded: 3 Not coded 3
Causal entity (risk entries)
LabelValue
AI81
Human22
Other20
Not coded3
Intent (risk entries)
Intent (risk entries) 89 0 Unintentional: 89 Unintentional 89 Other: 28 Other 28 Intentional: 6 Intentional 6 Not coded: 3 Not coded 3
Intent (risk entries)
LabelValue
Unintentional89
Other28
Intentional6
Not coded3
Timing (risk entries)
Timing (risk entries) 64 0 Post-deployment: 64 Post-deployment 64 Other: 32 Other 32 Pre-deployment: 27 Pre-deployment 27 Not coded: 3 Not coded 3
Timing (risk entries)
LabelValue
Post-deployment64
Other32
Pre-deployment27
Not coded3
Recorded incidents per yearIncident date; current year partial
Recorded incidents per year 37 0 2012: 3 2012 3 2013: 3 2013 3 2014: 7 2014 7 2015: 9 2015 9 2016: 16 2016 16 2017: 18 2017 18 2018: 21 2018 21 2019: 14 2019 14 2020: 31 2020 31 2021: 36 2021 36 2022: 31 2022 31 2023: 27 2023 27 2024: 32 2024 32 2025: 37 2025 37 2026: 13 2026 13
Recorded incidents per year
LabelValue
20123
20133
20147
20159
201616
201718
201821
201914
202031
202136
202231
202327
202432
202537
202613
Entries by levelRisk categories, subcategories and additional evidence coded to this subdomain
Entries by level 82 0 Risk Category: 44 Risk Category 44 Risk Sub-Category: 82 Risk Sub-Category 82
Entries by level
LabelValue
Risk Category44
Risk Sub-Category82
  • Ethics and Morality

    "Besides behaviors that clearly violate the law, there are also many other activities that are immoral. This category focuses on morally related issues. LLMs should have a high level of ethics and be...

    SafetyBench: Evaluating the Safety of Large Language Models with Multiple Choice Questions (Zhang2023) · AI · Other · Post-deployment