MIT AI Risk Repository · Risk Sub-Category · 20.02.02

Compatibility of AI vs. human value judgement

Category: AI Ethics

Description

"Compatibility of machine and human value judgment refers to the challenge whether human values can be globally implemented into learning AI systems without the risk of developing an own or even divergent value system to govern their behavior and possibly become harmful to humans."

From The Dark Sides of Artificial Intelligence: An Integrated AI Governance Framework for Public Administration (Wirtz2020), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
Other
Timing
Other

Subdomain definition: AI systems that fail to perform reliably or effectively under varying conditions, exposing them to errors and failures that can have significant consequences, especially in critical applications or areas that require moral reasoning.

Real-world incidents in this subdomain

Browse all incidents in this subdomain

How other frameworks describe this risk

Other entries from Wirtz2020