MIT AI Risk Repository

Browse AI risks

3 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

3 entries

  1. 67.04.03 · Risk Sub-Category

    Loss of control

    Capabilities that could be used to reduce human control - Manipulation

    "There is evidence that language models tend to respond as though they share the user’s stated views, and larger models do this more than smaller ones.276 The ability to predict people’s views and generate text that they will endorse could be useful for manipulation."

    From Capabilities and Risks from Frontier AI (DSIT2023)

  2. 67.04.04 · Risk Sub-Category

    Loss of control

    Capabilities that could be used to reduce human control - Cyber offence

    "Instead of - or in addition to - manipulating humans, AI systems could acquire influence by exploiting vulnerabilities in computer systems. Offensive cyber capabilities could allow AI systems to gain access to money, computing resources, and critical infrastructure. As discussed earlier in this report, frontier AI is already lowering the barrier for threat actors and future AI agents may be able to execute cyber attacks autonomously.":

    From Capabilities and Risks from Frontier AI (DSIT2023)

  3. 67.04.05 · Risk Sub-Category

    Loss of control

    Capabilities that could be used to reduce human control - Autonomous replication and adaptation

    "Controlling AI systems could become much harder if they could autonomously persist, replicate, and adapt in cyberspace. No current AI systems have this capability, but recent research found that frontier AI agents can perform some relevant tasks.279"

    From Capabilities and Risks from Frontier AI (DSIT2023)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.