MIT AI Risk Repository · Risk Category · 04.07.00

Malicious Use and Unleashing AI Agents

Description

LMs, due to their remarkable capabilities, carry the same potential for malice as other technological products. For instance, they may be used in information warfare to generate deceptive information or unlawful content, thereby having a significant impact on individuals and society. As current LMs are increasingly built as agents to accomplish user objectives, they may disregard the moral and safety guidelines if operating without adequate supervision. Instead, they may execute user commands mechanically without considering the potential damage. They might interact unpredictably with humans a

From Towards Safer Generative Language Models: A Survey on Safety Risks, Evaluations, and Improvements (Deng2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Subdomain
4.0
Causal entity
Other

How other frameworks describe this risk

Other entries from Deng2023