MIT AI Risk Repository

Browse AI risks

6 risk entries extracted from 74 frameworks, coded by domain, subdomain, causal entity, intent and timing. Filter, then export the current selection with its licence and citation attached.

Reset Also filtered by framework Gabriel2024 ×

6 entries

  1. 24.11.00 · Risk Category

    Misinformation risks

    "The rapid integration of AI systems with advanced capabilities, such as greater autonomy, content generation, memorisation and planning skills (see Chapter 4) into personalised assistants also raises new and more specific challenges related to misinformation, disinformation and the broader integrity of our information environment. "

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  2. 24.06.01 · Risk Sub-Category

    Appropriate Relationships

    Causing direct emotional or physical harm to users

    AI assistants could cause direct emotional or physical harm to users by generating disturbing content or by providing bad advice. "Indeed, even though there is ongoing research to ensure that outputs of conversational agents are safe (Glaese et al., 2022), there is always the possibility of failure modes occurring. An AI assistant may produce disturbing and offensive language, for example, in response to a user disclosing intimate information about themselves that they have not felt comfortable sharing with anyone else. It may offer bad advice by providing factually incorrect information (e.g.

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  3. 24.11.01 · Risk Sub-Category

    Misinformation risks

    Entrenched viewpoints and reduced political efficacy

    "Design choices such as greater personalisation of AI assistants and efforts to align them with human preferences could also reinforce people’s pre-existing biases and entrench specific ideologies. Increasingly agentic AI assistants trained using techniques such as reinforcement learning from human feedback (RLHF) and with the ability to access and analyse users’ behavioural data, for example, may learn to tailor their responses to users’ preferences and feedback. In doing so, these systems could end up producing partial or ideologically biased statements in an attempt to conform to user expec

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  4. 24.11.02 · Risk Sub-Category

    Misinformation risks

    Degraded and homogenised information environments

    "Beyond this, the widespread adoption of advanced AI assistants for content generation could have a number of negative consequences for our shared information ecosystem. One concern is that it could result in a degradation of the quality of the information available online. Researchers have already observed an uptick in the amount of audiovisual misinformation, elaborate scams and fake websites created using generative AI tools (Hanley and Durumeric, 2023). As more and more people turn to AI assistants to autonomously create and disseminate information to public audiences at scale, it may beco

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  5. 24.11.05 · Risk Sub-Category

    Misinformation risks

    Entrenching specific ideologies

    "AI assistants may provide ideologically biased or otherwise partial information in attempting to align to user expectations. In doing so, AI assistants may reinforce people’s pre-existing biases and compromise productive political debate."

    From The Ethics of Advanced AI Assistants (Gabriel2024)

  6. 24.11.06 · Risk Sub-Category

    Misinformation risks

    Eroding trust and undermining shared knowledge

    "AI assistants may contribute to the spread of large quantities of factually inaccurate and misleading content, with negative consequences for societal trust in information sources and institutions, as individuals increasingly struggle to discern truth from falsehood."

    From The Ethics of Advanced AI Assistants (Gabriel2024)

Informational only, not legal advice. Verify every claim against the linked official sources and consult qualified counsel before acting.