AI Risk Atlas Prototype/DemoUnofficial independent experiment. Not an official xAI product. Scores can be wrong.

Back to mitigations
Proposed

Adversarial test sets built from civilian lookalikes

Evaluate on the distribution that gets people killed, not on the distribution that wins a demo.

No compiled flow or public database names this control yet.

Who should own it
Defence labs
Governments
How quickly it can land
Months
A planned program, one to two quarters.
Expedited implementation
2 months
60 calendar days with a crash team
Normal implementation
7 months
210 calendar days as a planned program

Risk this mitigates

15
Battlefield AI inventing targets

Given that military systems are being fielded that can propose or prosecute targets, and warnings already exist that they invent intent, there is a possibility of a false positive becoming a kinetic event resulting in civilian casualties, unlawful strikes, and rapid escalation between states.

Residual composite 15 · still above the threshold

Effect if implemented

Applied to every failure scenario on that risk, then re-ranked. Axes are clamped at 1.

Likelihood
1
Consequence
0
Urgency
0