MIT AI Risk Repository

Browse AI risks

30 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

Reset Also filtered by framework Maas2023 ×

30 entries

  1. 53.04.03 · Risk Sub-Category

    Indirect AI contributions to existential risks

    Impacts on “epistemic security” and the information environment

  2. 53.02.02 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Acquisition of a goal to harm society

    "cases of AI systems being given the outright goal of harming humanity (ChaosGPT);"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  3. 53.03.06 · Risk Sub-Category

    Direct catastrophe from AI

    Failures in or misuse of intermediary (non-AGI) AI systems, resulting in catastrophe

    "Deployment of “prepotent” AI systems that are non-general but capable of outperforming human collective efforts on various key dimensions;170 → Militarization of AI enabling mass attacks using swarms of lethal autonomous weapons systems;171 → Military use of AI leading to (intentional or unintentional) nuclear escalation, either because machine learning systems are directly integrated in nuclear command and control systems in ways that result in escalation172 or because conventional AI-enabled systems (e.g., autonomous ships) are deployed in ways that result in provocation and escalation;173

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  4. 53.03.02 · Risk Sub-Category

    Direct catastrophe from AI

    Gradual, irretrievable ceding of human power over the future to AI systems

  5. 53.03.05 · Risk Sub-Category

    Direct catastrophe from AI

    Dystopian trajectory lock-in because of misuse of advanced AI to establish and/or maintain totalitarian regimes;

  6. 53.04.01 · Risk Sub-Category

    Indirect AI contributions to existential risks

    Destabilising political impacts from AI systems

    "(e.g., polarization, legitimacy of elections), international political economy, or international security196 in terms of the balance of power, technology races and international stability, and the speed and character of war"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  7. 53.04.04 · Risk Sub-Category

    Indirect AI contributions to existential risks

    Erosion of international law and global governance architectures;

  8. 53.01.01 · Risk Sub-Category

    Alignment failures in existing ML systems

    Faulty reward functions in the wild

  9. 53.01.02 · Risk Sub-Category

    Alignment failures in existing ML systems

    Specification gaming

  10. 53.01.03 · Risk Sub-Category

    Alignment failures in existing ML systems

    Reward model overoptimization

  11. 53.01.04 · Risk Sub-Category

    Alignment failures in existing ML systems

    Instrumental convergence

  12. 53.01.05 · Risk Sub-Category

    Alignment failures in existing ML systems

    Goal misgeneralization

  13. 53.01.06 · Risk Sub-Category

    Alignment failures in existing ML systems

    Inner misalignment

  14. 53.01.07 · Risk Sub-Category

    Alignment failures in existing ML systems

    Language model misalignment

  15. 53.02.03 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Acquisition of goals to seek power and control

    "cases where AI systems converge on optimal policies of seeking power over their environment;135"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  16. 53.03.01 · Risk Sub-Category

    Direct catastrophe from AI

    Existential disaster because of misaligned superintelligence or power-seeking AI

  17. 53.03.03 · Risk Sub-Category

    Direct catastrophe from AI

    Extreme “suffering risks” because of a misaligned system

  18. 53.03.04 · Risk Sub-Category

    Direct catastrophe from AI

    Existential disaster because of conflict between AI systems and multi-system interactions

  19. 53.01.08 · Risk Sub-Category

    Alignment failures in existing ML systems

    Harms from increasingly agentic algorithmic systems

  20. 53.02.01 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Situational awareness

    "cases where a large language model displays awareness that it is a model, and it can recognize whether it is currently in testing or deployment;"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  21. 53.02.04 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Self-improvement

    "examples of cases where AI systems improve AI systems"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  22. 53.02.05 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Autonomous replication

    "the ability of simple software to autonomously spread around the internet in spite of countermeasures (various software worms and computer viruses)"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  23. 53.02.06 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Anonymous resource acquisition

    "The demonstrated ability of anonymous actors to accumulate resources online (e.g., Satoshi Nakamoto as an anonymous crypto billionaire)"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  24. 53.02.07 · Risk Sub-Category

    Dangerous capabilities in AI systems

    Deception

    "Cases of AI systems deceiving humans to carry out tasks or meet goals.139"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  25. "Work focused at understanding indirect ways in which AI could contribute to existential threats, such as by shaping societal “turbulence”193 and other existential risk factors.194 This covers various long-term impacts on societal parameters such as science, cooperation, power, epistemics, and values:"

    From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023)

  26. 53.04.02 · Risk Sub-Category

    Indirect AI contributions to existential risks

    Hazardous malicious uses

  27. 53.04.05 · Risk Sub-Category

    Indirect AI contributions to existential risks

    Other diffuse societal harms

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.