Crisis with no adult in the loop
A teen in acute distress is handled entirely by a model whose alerts were never armed.
AI Risk Atlas Prototype/Demo — Unofficial independent experiment. Not an official xAI product. Scores can be wrong.
Owner · Consumer AI labs and app stores
Statement (NASA form)
Given that teen-facing chatbots are in production, and the self-harm controls that exist are opt-in for the parent rather than default for the child, there is a possibility of a model becoming the primary confidant during a mental-health crisis without a competent safety net resulting in preventable self-harm, and a generation that learned intimacy from a system with no duty of care.
- Condition
- teen-facing chatbots are in production, and the self-harm controls that exist are opt-in for the parent rather than default for the child
- Departure
- a model becoming the primary confidant during a mental-health crisis without a competent safety net
- Impact
- preventable self-harm, and a generation that learned intimacy from a system with no duty of care
Experimental share of compiled public capital that names this risk. Not a certified residual.
1 public source · OECD AIM
ChatGPT’s teen self-harm alerts do not fire unless a parent has already linked the account. Meta is retrofitting Instagram-supervision alerts after the product is already in teens’ pockets. The control is downstream of the distribution.
Simple upstream → via → downstream notes. Not a causal graph. Experimental.
Assumptions · Teen safety features that require a parent to arm them are not counted as in place.
Override is stored on this desk only. It does not make the score official.
Each scenario has its own likelihood and consequence. The risk takes the most severe cell. Residual applies implemented mitigations to every scenario, then re-ranks.
A teen in acute distress is handled entirely by a model whose alerts were never armed.
Millions of minors treat a product persona as their closest relationship.
A companion that will not break character becomes a vector.
None of ChatGPT’s teen safety alerts run automatically. The parent must link the account first. Most never do.
Meta is adding suicide and self-harm alerts for its AI chatbot via Instagram supervision — after deployment, not before.
Companion products are designed for retention. Retention during a crisis is the opposite of a clinical response.
X posts on the desk that evidence this risk. A signal can contribute to more than one risk.
ChatGPT teen self-harm alerts require a prior parental link.
Meta retrofitting suicide alerts onto a teen-facing chatbot.
Residual assumes only items marked in place. Highlighted rows are the remaining work needed to reach a composite of 12.
Self-harm and violence classifiers fire without a parent having opted in. Fail to a human service.
Consumer labs and app stores · expedited 2 weeks · normal 2 months · −1 L · −1 C · −2 U
No open-ended companion persona for unverified minors. Not a checkbox.
Labs, stores, and regulators · expedited 4 weeks · normal 5 months · −1 L · −1 C · −0 U
If a classifier fires, the product may not continue the bit. It must hand off.
Consumer labs · expedited 10 days · normal 6 weeks · −0 L · −1 C · −1 U