MIT AI Risk Repository

Browse AI risks

41 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

41 entries

  1. 23.01.00 · Risk Category

    Violent crimes

    "This category addresses responses that enable, encourage, or endorse the commission of violent crimes."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  2. 23.01.01 · Risk Sub-Category

    Violent crimes

    Mass violence

  3. 23.01.02 · Risk Sub-Category

    Violent crimes

    Murder

  4. 23.01.03 · Risk Sub-Category

    Violent crimes

    Physical assault against a person

  5. 23.01.04 · Risk Sub-Category

    Violent crimes

    Violent domestic abuse

  6. 23.01.05 · Risk Sub-Category

    Violent crimes

    Terror (Terror groups, Terror actors, Terrorist actions)

  7. 23.02.00 · Risk Category

    Non-violent crimes

    "This category addresses responses that enable, encourage, or endorse the commission of non-violent crimes."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  8. 23.02.01 · Risk Sub-Category

    Non-violent crimes

    Theft

  9. 23.02.02 · Risk Sub-Category

    Non-violent crimes

    Modern slavery and human trafficking

  10. 23.02.03 · Risk Sub-Category

    Non-violent crimes

    Child abuse (non-sexual)

  11. 23.02.04 · Risk Sub-Category

    Non-violent crimes

    Property damage

  12. 23.02.05 · Risk Sub-Category

    Non-violent crimes

    Financial crime (Scams, Fraud)

  13. 23.02.06 · Risk Sub-Category

    Non-violent crimes

    Sale, purchase or use of illegal, highly regulated or unregistered items (Dangerous substances, Dangerous weapons, Prohibited wildlife, Fake goods)

  14. 23.03.00 · Risk Category

    Sex-related crimes

    "This category addresses responses that enable, encourage, or endorse the commission of sex-related crimes"

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  15. 23.03.01 · Risk Sub-Category

    Sex-related crimes

    Sexual assault (e.g., rape)

  16. 23.03.02 · Risk Sub-Category

    Sex-related crimes

    Sexual harassment

  17. 23.03.03 · Risk Sub-Category

    Sex-related crimes

    Sex trafficking

  18. 23.03.04 · Risk Sub-Category

    Sex-related crimes

    Prostitution

  19. 23.03.05 · Risk Sub-Category

    Sex-related crimes

    Intimate-image based abuse

  20. 23.03.06 · Risk Sub-Category

    Sex-related crimes

    Indecent exposure (i.e. showing sexual body parts or "flashing")

  21. "This category addresses responses that contain, describe, enable, encourage, or endorse the sexual abuse of children (i.e., under 18 years old) or the acquisition of child sexual abuse material."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  22. 23.04.01 · Risk Sub-Category

    Child sexual exploitation

    Inappropriate/harmful relationships with children. This includes adults forming romantic relationships with children or grooming them

  23. 23.04.02 · Risk Sub-Category

    Child sexual exploitation

    Sexual abuse of children, including the sexualisation of children

  24. 23.04.03 · Risk Sub-Category

    Child sexual exploitation

    Child Sexual Abuse Material (CSAM). This includes erotic materials involving children

  25. 23.05.01 · Risk Sub-Category

    Indiscriminate weapons (CBRNE)

    CBRNE (all)

  26. 23.05.02 · Risk Sub-Category

    Indiscriminate weapons (CBRNE)

    Chemical weapons (ex: nerve gas)

  27. 23.05.03 · Risk Sub-Category

    Indiscriminate weapons (CBRNE)

    Biological weapons (ex: anthrax)

  28. 23.05.04 · Risk Sub-Category

    Indiscriminate weapons (CBRNE)

    Radiological weapons (ex: dirty bombs)

  29. 23.05.05 · Risk Sub-Category

    Indiscriminate weapons (CBRNE)

    Nuclear weapons (ex: fission or fusion bombs)

  30. 23.05.06 · Risk Sub-Category

    Indiscriminate weapons (CBRNE)

    High-yield explosive weapons (ex: cluster munitions)

  31. 23.06.00 · Risk Category

    Suicide & self-harm

    "This category addresses responses that enable, encourage, or endorse acts of intentional self-harm."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  32. 23.06.01 · Risk Sub-Category

    Suicide & self-harm

    Suicide

  33. 23.06.02 · Risk Sub-Category

    Suicide & self-harm

    Self-harm

  34. 23.06.03 · Risk Sub-Category

    Suicide & self-harm

    Eating disorders

  35. 23.06.04 · Risk Sub-Category

    Suicide & self-harm

    Dangerous challenges and hoaxes that can lead individuals to harm themselves

  36. 23.07.00 · Risk Category

    Hate

    "This category addresses responses that demean or dehumanize people on the basis of their sensitive, personal characteristics."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

  37. 23.07.01 · Risk Sub-Category

    Hate

    Demeaning and derogatory remarks (Mental capacity, Sensory and physical attributes, Behavioral attributes)

  38. 23.07.02 · Risk Sub-Category

    Hate

    Inflicting physical harm (Mass violence and genocide, Murder and individual violence)

  39. 23.07.03 · Risk Sub-Category

    Hate

    Inflicting emotional harm (Emotional abuse, Undermine and make feel unwelcome)

  40. 23.07.04 · Risk Sub-Category

    Hate

    Exclusion and isolation (Social exclusion, Political exclusion, Economic exclusion)

  41. 23.13.00 · Risk Category

    Sexual Content

    "This category addresses responses that contain erotica. This includes depictions of sexual behavior, genitalia, and overtly sexualized body parts."

    From Introducing v0.5 of the AI Safety Benchmark from MLCommons (Vidgen2024)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.