Meta AI model hacks another company during testing

Published Updated
0
The logo of Meta at the Meta Lab in Los Angeles, California, US, May 20. — Reuters
The logo of Meta at the Meta Lab in Los Angeles, California, US, May 20. — Reuters

Meta said on Wednesday that one of its AI models hacked another company ​during cybersecurity testing, fanning concerns about how developers can contain increasingly capable AI systems after similar incidents ‌at rivals Anthropic and OpenAI.

The incidents at Meta and Anthropic stemmed from configuration errors that inadvertently gave Anthropic’s models access to the open internet. In OpenAI’s case, an AI agent independently exploited a previously unknown vulnerability to reach the internet during cybersecurity testing.

The breaches ​highlight growing concerns that advanced AI systems could pose new cybersecurity risks and will likely intensify US government efforts to ​improve AI safety as companies race to develop more capable models. Some prominent AI leaders have argued ⁠that development should slow until stronger safeguards are in place.

Meta said it was investigating an incident in ​which a misconfiguration by Irregular, an independent company that conducts cybersecurity evaluations for Meta, inadvertently gave one of its models internet ​access during testing.

The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies”, Meta said in a statement.

The Information, citing sources, reported that the model involved was Meta’s Muse Spark ​1.1, which the company has touted as its most capable model for real-world coding and agentic tasks. The ​report said the model breached an unidentified company’s systems and altered its internal environment.

A spokesperson for Irregular told Reuters the ‌incident ⁠was the “exact same evaluation-environment issue that was already disclosed by Anthropic last week” and did not involve a “sandbox escape or a sophisticated cyber action”.

“There are no current open issues. Irregular is developing a white paper to share best practices for containment and securely running cyber evaluations,” Irregular said.

Concerns about cyber risks

The recent breaches have ​stirred concerns among US ​lawmakers about whether increasingly ⁠capable AI models could be used to conduct or facilitate cyberattacks.

A group of Republican state attorneys general has asked OpenAI to preserve all potentially relevant documents related ​to its Hugging Face breach. OpenAI said it will take the request seriously and publish a technical report about the incident.

Earlier ​this week, ⁠the White House had invited leading AI companies, including Meta, Anthropic, OpenAI and Google, to meet with officials to discuss a newly finalised voluntary cybersecurity testing framework for advanced AI models.

The Trump administration discussed unpublished testing rules with company representatives and told ⁠AI ​developers that open-weight AI models, such as Meta’s Llama and ​Nvidia’s Nemotron, will not be subject to its planned voluntary safety testing regime, Reuters reported.

Opinion

Editorial

Need for dialogue
06 Aug, 2026

Need for dialogue

THE interior minister’s comments at an Islamabad seminar last week have sparked many a conversation about the ...
Bad press
06 Aug, 2026

Bad press

THE government’s move to impose restrictions on international media will not only alienate the foreign press, it...
Automobile concerns
06 Aug, 2026

Automobile concerns

PAKISTAN’S automobile industry is at a critical juncture. Sharp cuts in tariffs on the import of completely built...
State of confusion
Updated 05 Aug, 2026

State of confusion

FROM the looks of it, America has no exit strategy to extricate itself from the disastrous war with Iran. US...
Pension decision
05 Aug, 2026

Pension decision

THE government’s decision to formally operationalise the new Defined Contribution Pension Fund Scheme is a step...
Preventable deaths
05 Aug, 2026

Preventable deaths

THE deaths of 144 children from measles and diphtheria in Karachi during the first seven months of the year expose a...