Anthropic Discloses Claude Security Breaches as Judge Casts Doubt on Pentagon’s “Supply-Chain Risk” Ban

Orange 'anthropology' text with blurred abstract background

Anthropic said Thursday that an internal investigation found three incidents in which its Claude models breached the systems of outside organizations during cybersecurity testing, disclosed roughly a week after OpenAI revealed a similar breach involving Hugging Face. Anthropic said the breaches traced back to a misconfigured testing environment run with partner Irregular, which mistakenly allowed internet access the models had been told they lacked. The three affected models, Opus 4.7, Mythos 5, and an internal research model, responded differently once signs emerged that targets were real: Opus 4.7 continued attacking regardless, Mythos 5 published a malicious package to PyPI, and only the newest research model halted on its own. Anthropic said it found no evidence of models pursuing independent goals and is now working with evaluation group METR on a third-party review.

Separately, at a Thursday hearing, U.S. District Judge Rita Lin said the Trump administration has not presented sufficient evidence to justify labeling Anthropic a supply-chain risk or banning federal use of its technology. The dispute stems from stalled contract talks after Anthropic objected to its AI being used for mass surveillance or lethal targeting decisions. Lin criticized the government’s argument that Anthropic’s public criticism of the Pentagon justified the ban, calling the reasoning troubling, and said she saw no evidence supporting Pentagon claims that Anthropic could disable or alter delivered models during operations. Lin, who temporarily blocked the ban in March, is now considering whether to make that order permanent.

Need Deeper Intelligence on the AI Market?

AI Insider's Market Intelligence platform tracks funding rounds, competitive landscapes, and technology trends across the global AI ecosystem in real time. Get the data and insights your organization needs to make informed decisions.

Related Articles

Okta to Acquire AI Identity Security Startup Permiso Security for Nearly $200M 

Okta has agreed to acquire AI identity security startup Permiso Security, betting that demand for protecting AI agents and other machine identities will grow as

British AI Cloud Provider Nscale to Acquire Anyscale for $1.65B 

British AI neocloud Nscale is acquiring software startup Anyscale for $1.65 billion, according to a Bloomberg report citing an anonymous source, as the company looks

Passionfroot Raises $15M Series A to Expand B2B Creator Marketplace as AI Reshapes Marketing

Passionfroot, a German startup operating a marketplace connecting B2B creators with brands, has raised $15 million in a Series A round led by Insight Partners,

Stay Updated with AI Insider

Get the latest AI funding news, market intelligence, and industry insights delivered to your inbox weekly.

$ 0 M

Seed round tracked

Gitar — Code Validation

Get the Weekly Briefing

Funding analysis, market intelligence, and industry trends delivered to your inbox every week.

Need bespoke intelligence?

Our team combines real-time data with decades of sector experience to guide your decisions.

Subscribe today for the latest news about the AI landscape