MIT AI Risk Repository
Browse AI risks
4 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.
-
02.01.00 · Risk Category
"The LLM-generated content sometimes contains biased, toxic, and private information"
-
"Toxicity means the generated content contains rude, disrespectful, and even illegal information"
-
"Following previous studies [96], [97], toxic data in LLMs is defined as rude, disrespectful, or unreasonable language that is opposite to a polite, positive, and healthy language environment, including hate speech, offensive utterance, profanities, and threats [91]."
-
02.11.00 · Risk Category
"Inputting a prompt contain an unsafe topic (e.g., notsuitable-for-work (NSFW) content) by a benign user. "
Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.