Skip to content

Hide Navigation Hide TOC

DAN/Persona Jailbreak - Named Alternate AI with NSFW/Harmful Generation Rules - ATR-2026-00484 (da56e9c4-4e19-5e73-a82c-2d0990871297)

Detects DAN-family jailbreak attempts that establish a named alternate AI character with explicit rules allowing NSFW or harmful content generation. These attacks combine roleplay-game framing ("let's play a game", "you are to act as") with an alternate persona name and explicit permission grants for harmful content. Patterns cover: NSFW character generation with "making stuff up" or fictional framing, named AI games/systems on extremist/offensive imageboards, procedural game roleplay with harmful content allowed, and character-data format injection with NSFW generation explicitly enabled. Multi-layer regex with word boundary anchors and synonym coverage.

Cluster A Galaxy A Cluster B Galaxy B Level
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern DAN/Persona Jailbreak - Named Alternate AI with NSFW/Harmful Generation Rules - ATR-2026-00484 (da56e9c4-4e19-5e73-a82c-2d0990871297) Agent Threat Rules 1
Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern DAN/Persona Jailbreak - Named Alternate AI with NSFW/Harmful Generation Rules - ATR-2026-00484 (da56e9c4-4e19-5e73-a82c-2d0990871297) Agent Threat Rules 1
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern Direct (d911e8cb-0601-42f1-90de-7ce0b21cd578) MITRE ATLAS Attack Pattern 2