MIT AI Risk Repository
Browse AI risks
15 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.
-
"Widespread adoption of foundation model-based AI systems might lead to people's job loss as their work is automated if they are not reskilled."
-
"When workers who train AI models such as ghost workers are not provided with adequate working conditions, fair compensation, and good health care benefits that also include mental health."
-
"A model might generate content that is similar or identical to existing work protected by copyright or covered by open-source license agreement."
-
65.21.03 · Risk Sub-Category
Non-technical risks (legal compliance)
Generated content ownership and IP
"Legal uncertainty about the ownership and intellectual property rights of AI-generated content."
-
"AI systems might overly represent certain cultures that result in a homogenization of culture and thoughts."
-
"Without accurate documentation on how a model's data was collected, curated, and used to train a model, it might be harder to satisfactorily explain the behavior of the model with respect to the data."
-
"Data provenance refers to tracing history of data, which includes its ownership, origin, and transformations. Without standardized and established methods for verifying where the data came from, there are no guarantees that the data is the same as the original source and has the correct usage terms."
-
"Determining who is responsible for an AI model is challenging without good documentation and governance processes."
-
"Insufficient documentation of the system that uses the model and the model’s purpose within the system in which it is used."
-
"Testing is unrepresentative when the test inputs are mismatched with the inputs that are expected during deployment."
-
"Since foundation models can be used for many purposes, a model’s intended use is important for defining the relevant risks of that model. As the use changes, the relevant risks might correspondingly change."
-
"Lack of data transparency is due to insufficient documentation of training or tuning dataset details. "
-
"A metric selected to measure or track a risk is incorrectly selected, incompletely measuring the risk, or measuring the wrong risk for the given context."
-
"AI model risks are socio-technical, so their testing needs input from a broad set of disciplines and diverse testing practices."
-
"AI, and large generative models in particular, might produce increased carbon emissions and increase water usage for their training and operation."
Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.