AI Risk Atlas Prototype/DemoUnofficial independent experiment. Not an official xAI product. Scores can be wrong.

Back to mitigations
In progress

Model Hardware Standard with safety limits on physical agents

Anthropic/Janelia MHS is a shared driver for lab and factory gear. Do not open-source it until physical-safety evals exist: discoverability without interlocks is a new affordance, not a control.

Who should own it
Labs and hardware vendors
Frontier labs
How quickly it can land
Months
A planned program, one to two quarters.
Expedited implementation
2 months
60 calendar days with a crash team
Normal implementation
8 months
240 calendar days as a planned program

Risk this mitigates

15
Battlefield AI inventing targets

Given that military systems are being fielded that can propose or prosecute targets, and warnings already exist that they invent intent, there is a possibility of a false positive becoming a kinetic event resulting in civilian casualties, unlawful strikes, and rapid escalation between states.

Residual composite 15 · still above the threshold

Effect if implemented

Applied to every failure scenario on that risk, then re-ranked. Axes are clamped at 1.

Likelihood
1
Consequence
1
Urgency
0