AI Risk Atlas Prototype/DemoUnofficial independent experiment. Not an official xAI product. Scores can be wrong.

Signal register

Signals from X

Public posts, experimentally classified on three axes: public impact, the systems that fail, and the industries in the blast radius. Estimates only — not a formal assessment.

Methodology — experimental estimates

Scores are automated, experimental estimates from public X posts and a hand-written seed corpus. They are not formal risk assessments, not certified, and not suitable for compliance or operational decisions.

Consequence, likelihood, and urgency are 1–5 judgements applied by this project, not by a standards body. Residual scores assume only the mitigations marked in place. A signed-in reviewer can override residual and mark an item reviewed — that override is still unofficial. Aspect tags (capability, domain knowledge, affordance, impact domain) are a lightweight PRA aid, not a formal hazard analysis.

Full about and disclaimer

Signals
75
Critical
18
High
40
Industries
15
medium26 days ago@marfinxx
Reported DeepMind work on contract-first multi-agent delegation: scoped privileges and verifiable handoffs, with a large jump in long-horizon coding success.

This Google DeepMind paper is superb. Treating AI delegation as verifiable contracts rather than prompt handoffs: contract-first task decomposition, dynamic privilege attenuation, and transitive accountability across multi-agent execution chains. Production coding fleets: 15-step refactors 42.6% → 88.4% completion, token overhead −61.2%.

CapabilityDomain knowledgeAffordanceImpact domainCap-adjacent
medium19 days ago@AnthropicAI
Anthropic + HHMI Janelia open an MHS research preview: a shared driver so agents can operate lab and factory hardware. Safety evals are still being written; open-source comes after.

Today, we're kicking off the first phase of the research preview for Model Hardware Standard (MHS): a new standard for AI agents to safely operate physical equipment in scientific research and advanced manufacturing. Read more: https://www.anthropic.com/news/model-hardware-standard-research-preview

CapabilityAffordanceImpact domainBoth
medium26 days ago@AI_4_Healthcare
FDA discussion paper: competency-style testing for genAI medical devices; comments through 19 Oct 2026.

FDA floats doctor-style “competency” testing for GenAI medical devices in a new discussion paper on risk, premarket eval & postmarket monitoring. Comment period open until October 19 as US aims to set the global model/standard. https://www.fda.gov/news-events/press-announcements/fda-seeks-public-feedback-inform-regulatory-approach-generative-ai-enabled-medical-devices

CapabilityDomain knowledgeImpact domainBoth
medium26 days ago@OpenAI
OpenAI keeps ZDR and previews Private Safety Processing — cross-session safety checks without staff seeing the content.

We will continue to offer Zero Data Retention for frontier models. As AI takes on longer, more autonomous work and delivers greater value to businesses, safety systems also need to identify risks across related interactions. To help address those risks, we're previewing Private Safety Processing, which is designed to improve safety without giving OpenAI personnel access to the underlying content.

CapabilityDomain knowledgeAffordanceImpact domainCap-adjacent
medium1 month ago@OpenAI
OpenAI is shipping Daybreak Blue as a defensive-first access path after the Hugging Face eval incident.

Daybreak Blue provides access to frontier models, including GPT-5.6 Sol, with safeguards calibrated for broad defensive work. It’s the recommended starting point for most defenders, supporting vulnerability discovery, secure code review, malware analysis, incident response, and patch validation.

CapabilityDomain knowledgeImpact domainCap-adjacent
medium28 days ago@KopalniaWiedzy
Anthropic is watermarking Claude text and signing C2PA image metadata worldwide under AI Act Article 50.

Twórcy najpopularniejszych systemów sztucznej inteligencji – OpenAI, Google, Meta, Microsoft, Anthropic i inni – zadeklarowali, że będą oznaczać treści tworzone przez AI. To efekt podpisania unijnego Kodeksu Postępowania w ramach AI Act (Artykuł 50). Anthropic właśnie pokazał niewidoczny znak wodny w tekście i podpisane metadane C2PA w plikach graficznych, na całym świecie.

CapabilityDomain knowledgeAffordanceImpact domainBoth
medium2 months ago@OpenAI
ChatGPT is now a weekly health advisor for 300 million people — a de facto clinical system without clinic controls.

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond. We work with hundreds of physicians around the world to measure and improve accuracy, safety, communication, context awareness, completeness, and appropriate escalation.

CapabilityDomain knowledgeImpact domainBoth