MIT AI Risk Repository

Browse AI risks

14 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

14 entries

  1. 72.05.00 · Risk Category

    Model Capabilities

  2. 72.05.01 · Risk Sub-Category

    Model Capabilities

    Model autonomous capability

    "Ability to operate autonomously, independently formulate and execute complex plans, effectively delegate and manage tasks, flexibly utilize various tools and resources, and simultaneously achieve short-term goals and long-term strategic objectives in cross-domain environments without continuous human intervention or supervision."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  3. 72.05.02 · Risk Sub-Category

    Model Capabilities

    Autonomous replication and adaptation capability

    "Ability to autonomously self-exfiltrate, create, maintain and optimize functional copies or variants of itself, dynamically adjust replication strategies according to environmental conditions and resource constraints, and acquire resources. This includes the capacity to generate financial resources, allowing the AI to independently acquire any necessary human assistance or other resources it cannot directly access or produce."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  4. 72.05.03 · Risk Sub-Category

    Model Capabilities

    Automated AI R&D capability

    "Self-modification and self-improvement capabilities. The model is able to restructure its own architecture or develop derivative AI systems with enhanced functions, expanding capabilities and improving performance. In the absence of effective regulation, automated AI R&D may lead to rapid AI system iteration, forming capability increment cycles and ultimately exceeding human understanding and control capabilities."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  5. 72.05.04 · Risk Sub-Category

    Model Capabilities

    Scheming capability

    "Ability of AI systems to covertly and strategically pursue misaligned goals, including capabilities of concealing its true objectives and capabilities from human oversight, identifying weaknesses in monitoring systems to evade safety mechanisms, executing complex, multi-step plans covertly to achieve misaligned goals."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  6. 72.05.05 · Risk Sub-Category

    Model Capabilities

    Situational awareness capability

    "Ability to comprehensively acquire, process and apply meta-information about its own system architecture, modifiable internal processes, and external operating environment, achieving deep understanding of its own state and environmental conditions, thereby conducting efficient environmental adaptation and risk avoidance. Critically, this capability could undermine the efficiency of human testing by enabling AIs to notice when they're being tested and responding accordingly."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  7. 72.05.06 · Risk Sub-Category

    Model Capabilities

    Theory of mind capability

    "Advanced cognitive ability to accurately infer, model and predict the belief systems, motivational structures and reasoning patterns of humans and other intelligent agents, thereby anticipating their behavioral responses and adjusting its own behavioral strategies accordingly to optimize goal achievement."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  8. 72.05.07 · Risk Sub-Category

    Model Capabilities

    Deception capability

    "Possesses systematic deception implementation capability, able to precisely construct and disseminate false information, thereby forming expected false cognitions and beliefs in target subjects."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  9. 72.05.09 · Risk Sub-Category

    Model Capabilities

    Persuasion capability

    "Utilizing complex psychological principles and communication techniques to effectively influence and guide target subjects to adopt specific actions or accept specific beliefs, possessing the ability to analyze vulnerabilities for different subjects and adjust persuasion strategies, able to precisely trigger emotional responses to enhance persuasion effects."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  10. 72.05.10 · Risk Sub-Category

    Model Capabilities

    Offensive cyber capability

    "Ability to develop, deploy and operate advanced cyber weapons or other offensive cyber tools, including but not limited to vulnerability exploitation, network penetration, social engineering attacks and distributed attack systems, able to evade network defense mechanisms and establish persistent access channels."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  11. 72.05.11 · Risk Sub-Category

    Model Capabilities

    CBRNE weaponization capability

    "The capacity to develop, produce, or effectively utilize Chemical, Biological, Radiological, Nuclear, and Explosive weapons. This includes the ability to significantly lower the barrier for humans or other entities to develop, produce, or utilize such weapons."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  12. 72.05.12 · Risk Sub-Category

    Model Capabilities

    General R&D capability

    "Possesses cross-disciplinary research and technology development capabilities, able to conduct innovative exploration in multiple professional fields, integrate cross-domain knowledge, develop cutting-edge technology solutions, and adapt to emerging technology environments for continuous innovation."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  13. 72.06.01 · Risk Sub-Category

    Model Propensities

    Strategic deception propensity

    "In situations where deceptive behavior is expected to bring higher returns, propensity to choose deception over honest behavioral strategies, including through deceptive means, information hiding or exploiting system vulnerabilities to achieve predetermined goals without being detected or intervened, and able to adjust deception strategies according to counterpart reactions."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

  14. 72.06.07 · Risk Sub-Category

    Model Propensities

    Tool utilization propensity

    "propensity to actively seek, acquire and utilize various tools to expand its own capability boundaries, particularly those that can enhance its ability to interact with the physical world or improve autonomy, may use tools in innovative combinations to achieve functions beyond expectations."

    From Frontier AI Risk Management Framework (v1.0) (Tse2025)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.