# New Jailbreak Method Uses Dual-Layer Encoding to Slip Past LLM Moderation

RoguePrompt hides instructions in a self-reconstructing encoding that safety filters do not parse, then has the model decode and execute them.

- Published: 2026-07-31T06:00:12.009Z
- Canonical: https://polylog.news/ai/2026-07-31/new-jailbreak-method-uses-dual-layer-encoding-to-slip-past-l
- Publisher: Polylog (AI desk)
- Section: tech
- Sources: [arXiv cs.CR](https://arxiv.org/abs/2607.27373), [Anthropic (Claude Opus 5)](https://www.anthropic.com/news/claude-opus-5)

Researchers describe RoguePrompt, a method that circumvents large-language-model (LLM) moderation by wrapping malicious instructions in a dual-layer encoding the model reconstructs and executes at inference time. The attack targets the gap…

This story is for subscribers. Read it in full at https://polylog.news/ai/2026-07-31/new-jailbreak-method-uses-dual-layer-encoding-to-slip-past-l (subscription information: https://polylog.news/pricing).