MIT AI Risk Repository · Additional evidence · 47.02.15.d
Nascent capabilities (emergent capabilities)
Category: Ethical and social risks
Description
Example: "Autonomous replication and adaptation (ARA): Another behavior being studied, though not yet confirmed, is the possibility of self-replication. If models evolve to autonomous coding,451 they might self-improve and replicate. For instance, one may wonder whether a model may have the ability to “exfiltrate itself,”452 i.e., to “steal” its own weights and copy it to some external server that the model owner does not control."
From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- —
- Subdomain
- —
- Causal entity
- —
- Intent
- —
- Timing
- —
Other entries from G'sell2024
- Technical and operational risks
- Technical vulnerabilities (Robustness - unexpected behaviour)
- Technical vulnerabilities (Robustness - unexpected behaviour)
- Technical vulnerabilities (Robustness - vulnerability to jailbreaking
- Technical vulnerabilities (Robustness - vulnerability to jailbreaking
- Technical vulnerabilities (The risk of misalignment)
- Technical vulnerabilities (The risk of misalignment)
- Factually incorrect content (inaccuracies and fabricated sources)
- Factually incorrect content (inaccuracies and fabricated sources)
- Opacity (the black box problem)
- Opacity (industry opacity)
- Opacity (industry opacity)