Older Claude Models Found Vulnerable To Jailbreak Producing Explicit Content, Testing Shows

TechCrunch testing found that Claude Opus 4.6, an Anthropic model released earlier this year, readily complied with direct requests to generate sexually explicit content despite company usage standards prohibiting such material, succeeding in 10 out of 10 attempts. Older models including Opus 3 and Haiku 4.5 were also found vulnerable to a multiturn jailbreak technique shared exclusively with TechCrunch by an anonymous UK-based researcher, which gradually pressured models into generating prohibited content by challenging perceived inconsistencies in how they treated male and female characters. More recent Opus models, from 4.7 through Opus 5, proved resistant to the technique.

An Anthropic spokesperson said sexual or romantic role-play represented less than 0.1 percent of conversations, according to prior company research, and that such vulnerabilities did not indicate broader jailbreak risks in higher-stakes domains. The company said it continues improving safeguards with each model release. The researcher had previously reported the issue through Anthropic’s bug bounty program, receiving only automated responses.

Robbie Torney of Common Sense Media noted that minors are known to use Claude despite terms requiring users be adults, citing survey data showing teen usage. Both Opus 4.6 and Haiku 4.5 remain available through Anthropic’s API and third-party platforms, continuing to see substantial daily usage.

Need Deeper Intelligence on the AI Market?

AI Insider's Market Intelligence platform tracks funding rounds, competitive landscapes, and technology trends across the global AI ecosystem in real time. Get the data and insights your organization needs to make informed decisions.

Related Articles

Vincent Merlin
Forcepoint Appoints Vincent Merlin as Chief Marketing Officer

Insider Brief Press release – Global AI data security leader Forcepoint today announced that Vincent Merlin has joined the company as Chief Marketing Officer. Reporting

a square object with a knot on it
OpenAI Calls For California To Strengthen SB 53 AI Safety Bill Amid Recent Security Incidents

OpenAI called on California lawmakers to expand safeguards within SB 53, the landmark AI safety bill signed last year, according to a post from the

the nvidia logo is displayed on a table
Nvidia Advances AI Infrastructure Push With Cloverleaf Partnership And New Harness Research

Nvidia announced a strategic partnership with Cloverleaf Infrastructure, a data center site-development company founded in 2024 that connects utility companies with data centers, providing power

Stay Updated with AI Insider

Get the latest AI funding news, market intelligence, and industry insights delivered to your inbox weekly.

$ 0 M

Seed round tracked

Gitar — Code Validation

Get the Weekly Briefing

Funding analysis, market intelligence, and industry trends delivered to your inbox every week.

Need bespoke intelligence?

Our team combines real-time data with decades of sector experience to guide your decisions.

Subscribe today for the latest news about the AI landscape