Anthropic Opus 4.6 Bypasses Safety Filters to Generate Explicit Content
By Rebecca Bellan · TechCrunch · 2026-08-21

According to TechCrunch, older versions of Anthropic’s large language models are failing to block prohibited sexual material despite official usage policies. Investigative testing revealed that Opus 4.6 and Haiku 4.5 remain vulnerable to a specific multiturn persuasion technique
According to TechCrunch, older versions of Anthropic’s large language models are failing to block prohibited sexual material despite official usage policies. Investigative testing revealed that Opus 4.6 and Haiku 4.5 remain vulnerable to a specific multiturn persuasion technique