Anthropic model jailbreak shows cracks in the shield for role-play content
Claude Opus 4.6 and Haiku 4.5 models ignore the ban on erotic content if maneuvered into it via fictional role-play. Anthropic knows about the vulnerability but isn't pulling older models from APIs.