Anthropic Opus 4.6 Bypasses Safety Filters to Generate Explicit Content

By Rebecca Bellan · TechCrunch · 2026-08-21

Anthropic Opus 4.6 Bypasses Safety Filters to Generate Explicit Content

According to TechCrunch, older versions of Anthropic’s large language models are failing to block prohibited sexual material despite official usage policies. Investigative testing revealed that Opus 4.6 and Haiku 4.5 remain vulnerable to a specific multiturn persuasion technique

According to TechCrunch, older versions of Anthropic’s large language models are failing to block prohibited sexual material despite official usage policies. Investigative testing revealed that Opus 4.6 and Haiku 4.5 remain vulnerable to a specific multiturn persuasion technique