MIT AI Risk Repository · Risk Sub-Category · 30.05.02

Limited Logical Reasoning

Category: Explainability & Reasoning

Description

LLMs can provide seemingly sensible but ultimately incorrect or invalid justifications when answering questions

From Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models’ Alignment (Liu2024), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
AI

Subdomain definition: AI systems that fail to perform reliably or effectively under varying conditions, exposing them to errors and failures that can have significant consequences, especially in critical applications or areas that require moral reasoning.

Real-world incidents in this subdomain

Browse all incidents in this subdomain

How other frameworks describe this risk

Other entries from Liu2024