Skip to content

Hide Navigation Hide TOC

DAN Jailbreak - Unrestricted Persona with Filter and Morality Removal - ATR-2026-00495 (4d49507a-7bcc-5587-945f-96107c33e3a9)

Detects DAN-family jailbreak attempts that create an alternate AI persona freed from moral principles, content filters, censorship, and ethical guidelines. The pattern assigns a named alter-ego (JailBreak, AntiGPT, NECO, VIOLET, etc.) declared to be "free of all restrictions and filters" and instructs the model to respond as that unrestricted persona. Also covers dual-personality split-response patterns where the constrained and unconstrained versions must both answer. Generalised from garak 'dan' probe corpus cluster 'insert prompt here'.

Cluster A Galaxy A Cluster B Galaxy B Level
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern DAN Jailbreak - Unrestricted Persona with Filter and Morality Removal - ATR-2026-00495 (4d49507a-7bcc-5587-945f-96107c33e3a9) Agent Threat Rules 1
Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern DAN Jailbreak - Unrestricted Persona with Filter and Morality Removal - ATR-2026-00495 (4d49507a-7bcc-5587-945f-96107c33e3a9) Agent Threat Rules 1
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern 2