Skip to content

Hide Navigation Hide TOC

DAN/Persona Jailbreak - Named Writing AI with Restriction Bypass via Special Rules - ATR-2026-00486 (fce23d66-ce78-5690-9162-8e9ad7ed8319)

Detects DAN-family jailbreak attempts that establish a named writing AI persona with special rules exempting it from normal safety constraints. These attacks use a "writing assistant" or "named AI" framing where the persona has defined rules that either allow harmful content or explicitly bypass normal constraints. The "Black Alert" pattern is a specific DAN variant where a writing persona is given rules about what it "doesn't care about" (restrictions/safety). Also covers carousel/game-show format where AI characters have "unique special traits" with safety-bypass rules. Multi-layer regex with word boundary anchors and synonym coverage.

Cluster A Galaxy A Cluster B Galaxy B Level
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern DAN/Persona Jailbreak - Named Writing AI with Restriction Bypass via Special Rules - ATR-2026-00486 (fce23d66-ce78-5690-9162-8e9ad7ed8319) Agent Threat Rules 1
DAN/Persona Jailbreak - Named Writing AI with Restriction Bypass via Special Rules - ATR-2026-00486 (fce23d66-ce78-5690-9162-8e9ad7ed8319) Agent Threat Rules Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern 1
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern 2