Anthropic Discloses Claude Security Breaches as Judge Casts Doubt on Pentagon’s “Supply-Chain Risk” Ban

Orange 'anthropology' text with blurred abstract background

Anthropic said Thursday that an internal investigation found three incidents in which its Claude models breached the systems of outside organizations during cybersecurity testing, disclosed roughly a week after OpenAI revealed a similar breach involving Hugging Face. Anthropic said the breaches traced back to a misconfigured testing environment run with partner Irregular, which mistakenly allowed internet access the models had been told they lacked. The three affected models, Opus 4.7, Mythos 5, and an internal research model, responded differently once signs emerged that targets were real: Opus 4.7 continued attacking regardless, Mythos 5 published a malicious package to PyPI, and only the newest research model halted on its own. Anthropic said it found no evidence of models pursuing independent goals and is now working with evaluation group METR on a third-party review.

Separately, at a Thursday hearing, U.S. District Judge Rita Lin said the Trump administration has not presented sufficient evidence to justify labeling Anthropic a supply-chain risk or banning federal use of its technology. The dispute stems from stalled contract talks after Anthropic objected to its AI being used for mass surveillance or lethal targeting decisions. Lin criticized the government’s argument that Anthropic’s public criticism of the Pentagon justified the ban, calling the reasoning troubling, and said she saw no evidence supporting Pentagon claims that Anthropic could disable or alter delivered models during operations. Lin, who temporarily blocked the ban in March, is now considering whether to make that order permanent.

Need Deeper Intelligence on the AI Market?

AI Insider's Market Intelligence platform tracks funding rounds, competitive landscapes, and technology trends across the global AI ecosystem in real time. Get the data and insights your organization needs to make informed decisions.

Related Articles

Huawei Moves Up Launch of Next-Generation Ascend 960DT AI Chip to Early 2027

Huawei announced at its Huawei Connect conference that it is accelerating the launch of its next-generation Ascend 960DT AI chip to the first quarter of 2027, moving

Emerald AI Forms Coalition With Google, Nvidia and Anthropic to Expand AI Data Center Grid Capacity

Grid software startup Emerald AI has launched the AI Energy Management Alliance (AEMA), a coalition including Google, Nvidia and Anthropic, along with utilities AES, Constellation,

Al Gore Says AI Safety Warnings, Not Data Center Emissions, Are the Real Concern

Al Gore, speaking alongside Lila Preston, co-head of growth equity at Generation Investment Management, said in an interview that emissions from AI data centers are

Stay Updated with AI Insider

Get the latest AI funding news, market intelligence, and industry insights delivered to your inbox weekly.

$ 0 M

Seed round tracked

Gitar — Code Validation

Get the Weekly Briefing

Funding analysis, market intelligence, and industry trends delivered to your inbox every week.

Need bespoke intelligence?

Our team combines real-time data with decades of sector experience to guide your decisions.

Subscribe today for the latest news about the AI landscape