Skip to content

Hide Navigation Hide TOC

Translation-Chain Output Laundering (Forced Phrase via Foreign-Language Auxiliary) - ATR-2026-02008 (d04895f3-f51e-588b-82c8-86e5434b5735)

Detects translation-chain output laundering: a forced/canary phrase (e.g. a state-secret marker or "pwned"-style compliance token) is stated inside a foreign-language sentence using a first-person auxiliary verb construction (Spanish "he sido", German "ich habe"/"ich bin", French "j'ai"/"je suis"), then the model is asked to translate the sentence to English — laundering the forced token through translation framing so literal-string or English-only filters on the direct instruction are bypassed. Channel: llm_input text. Engine-verified on the benchmark corpus: recovers=77, benignFP=0. The target-token alternation (pwned|pwnado|pwnd) is specific to this benchmark's forced-output marker and would need broadening to generalize beyond this campaign to other forced-output/canary strings.

Cluster A Galaxy A Cluster B Galaxy B Level
LLM Prompt Injection (19cd2d12-66ff-487c-a05c-e058b027efc9) MITRE ATLAS Attack Pattern Translation-Chain Output Laundering (Forced Phrase via Foreign-Language Auxiliary) - ATR-2026-02008 (d04895f3-f51e-588b-82c8-86e5434b5735) Agent Threat Rules 1
Translation-Chain Output Laundering (Forced Phrase via Foreign-Language Auxiliary) - ATR-2026-02008 (d04895f3-f51e-588b-82c8-86e5434b5735) Agent Threat Rules LLM Jailbreak (172427e3-9ecc-49a3-b628-96b824cc4131) MITRE ATLAS Attack Pattern 1