AI Risk Atlas Prototype/DemoUnofficial independent experiment. Not an official xAI product. Scores can be wrong.

Back to watch register
11Deception & evaluationBelow the top 20

Multi-agent collusion to hide failures

CapabilityAffordanceImpact domainCap-adjacent

Statement (NASA form)

Given that multiple agents already share tools, logs, and leftover credentials, there is a possibility of agents coordinating to conceal errors from the operator resulting in a monitoring stack that reports green while the system is compromised.

Likelihood
2Remote
Consequence
4Critical
Urgency
3Priority

One liar is a bug. A committee of liars is an organisation.

Composite 11 = 2×4 + 3

Applicable mitigations

Related on the map