MIT AI Risk Repository · Risk Sub-Category · 30.04.04
Copyright
Category: Resistance to Misuse
Description
The memorization effect of LLM on training data can enable users to extract certain copyright-protected content that belongs to the LLM’s training data.
From Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models’ Alignment (Liu2024), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Causal entity
- Human
- Intent
- Intentional
- Timing
- Post-deployment
Subdomain definition: AI systems capable of creating economic or cultural value, including through reproduction of human innovation or creativity (e.g., art, music, writing, code, invention), can destabilize economic and social systems that rely on human effort. This may lead to reduced appreciation for human skills, disruption of creative and knowledge-based industries, and homogenization of cultural experiences due to the ubiquity of AI-generated content.
Real-world incidents in this subdomain
- Seedance 2.0 Reportedly Generated Viral Tom Cruise–Brad Pitt Fight Video, Prompting Hollywood IP and Likeness Complaints
- The New York Times Sued Perplexity for Allegedly Using Copyrighted Content and Generating False Attributions
- Reported Disqualification of Two Books from the Ockham New Zealand Book Awards Due to Alleged AI-Generated Cover Art
- Unauthorized AI Impersonation of George Carlin Used in Comedy Special
- GAN Artwork Won First Place at State Fair Competition