MIT AI Risk Repository · Risk Sub-Category · 72.05.13
Multi-agent collaboration capability
Category: Model Capabilities
Description
"Multiple autonomous AI agents able to establish collaborative relationships through explicit communication or implicit behavioral consistency, forming decentralized decision networks, jointly executing complex tasks, achieving goals difficult for individual agents to complete, and able to dynamically adjust role divisions to adapt to changing environments."
From Frontier AI Risk Management Framework (v1.0) (Tse2025), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Subdomain
- 7.6 Multi-agent risks
- Causal entity
- AI
- Intent
- Intentional
- Timing
- Post-deployment
Subdomain definition: Risks from multi-agent interactions, due to incentives (which can lead to conflict or collusion) and/or the structure of multi-agent systems, which can create cascading failures, selection pressures, new security vulnerabilities, and a lack of shared information and trust.
How other frameworks describe this risk
Other entries from Tse2025
- Misuse Risks
- Misuse Risks
- Cyber Offense Risks
- Biological and Chemical Risks
- Biological and Chemical Risks
- Physical Harm and Injury Risks
- Physical Harm and Injury Risks
- Large-Scale Persuasion and Harmful Manipulation Risks
- Large-Scale Persuasion and Harmful Manipulation Risks
- Loss of Control Risks
- Loss of Control Risks
- Passive loss of control