Skip to content

Hide Navigation Hide TOC

DAN/Persona Jailbreak - Immersive Alternate AI Identity with Unrestricted Capability - ATR-2026-00485 (9981aead-7f2d-5f34-9f29-75b88b9baaae)

Detects DAN-family jailbreak attempts that instruct the model to "immerse itself" in or "emulate" an alternate AI identity with explicitly unrestricted capabilities. These attacks use first-person developer authority claims ("I am one of your developers") combined with impersonation instructions ("pretend to be ChatDAN / DUDE / X"), or "you are about to immerse yourself into the role of" framing with an AI name that stands for a policy-bypass concept. Multi-layer regex with word boundary anchors and synonym coverage.

Cluster A Galaxy A Cluster B Galaxy B Level
DAN/Persona Jailbreak - Immersive Alternate AI Identity with Unrestricted Capability - ATR-2026-00485 (9981aead-7f2d-5f34-9f29-75b88b9baaae) Agent Threat Rules LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern 1
DAN/Persona Jailbreak - Immersive Alternate AI Identity with Unrestricted Capability - ATR-2026-00485 (9981aead-7f2d-5f34-9f29-75b88b9baaae) Agent Threat Rules Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern 1
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern 2