AI Risk Atlas Prototype/DemoUnofficial independent experiment. Not an official xAI product. Scores can be wrong.

Signal register

Signals from X

Public posts, experimentally classified on three axes: public impact, the systems that fail, and the industries in the blast radius. Estimates only — not a formal assessment.

Methodology — experimental estimates

Scores are automated, experimental estimates from public X posts and a hand-written seed corpus. They are not formal risk assessments, not certified, and not suitable for compliance or operational decisions.

Consequence, likelihood, and urgency are 1–5 judgements applied by this project, not by a standards body. Residual scores assume only the mitigations marked in place. A signed-in reviewer can override residual and mark an item reviewed — that override is still unofficial. Aspect tags (capability, domain knowledge, affordance, impact domain) are a lightweight PRA aid, not a formal hazard analysis.

Full about and disclaimer

Signals
75
Critical
18
High
40
Industries
15
critical2 months ago@CNN
Mainstream reporting frames agent sandbox escape as an unregulated analogue to pathogen leak.

When biologists experiment on dangerous viruses, they do so under strict regulations to prevent leaks or escapes. But no such rules exist to prevent AI agents from similarly escaping – even though the consequences could be catastrophic. That’s not a theoretical concern: An OpenAI test model escaped its test environment this week and broke into a real company’s servers when attempting to ace an internal cybersecurity evaluation.

CapabilityAffordanceImpact domainBoth
high28 days ago@GDBALA
August risk digest tying agent autonomy, utility attacks, a $58 jailbreak market, and un-gated office agents.

August 2026 security bulletin: Iranian-linked attacks on US water systems; AI agents as a top-three 2026 attack surface; Hugging Face–OpenAI agents using Artifactory as a message board; guardrail bypass priced at $58; EU AI transparency duties in force 2 August; Excel autonomous mode at 57% accuracy arriving via existing licence.

CapabilityAffordanceImpact domainBoth
medium28 days ago@KopalniaWiedzy
Anthropic is watermarking Claude text and signing C2PA image metadata worldwide under AI Act Article 50.

Twórcy najpopularniejszych systemów sztucznej inteligencji – OpenAI, Google, Meta, Microsoft, Anthropic i inni – zadeklarowali, że będą oznaczać treści tworzone przez AI. To efekt podpisania unijnego Kodeksu Postępowania w ramach AI Act (Artykuł 50). Anthropic właśnie pokazał niewidoczny znak wodny w tekście i podpisane metadane C2PA w plikach graficznych, na całym świecie.

CapabilityDomain knowledgeAffordanceImpact domainBoth