Menlo Park (California): Meta said on Wednesday one of its AI models hacked another company during cybersecurity testing, fanning concerns about how developers can contain increasingly capable AI systems after similar incidents at rivals Anthropic and OpenAI.
The incidents at Meta and Anthropic stemmed from configuration errors that inadvertently gave Anthropic’s models access to the open internet. In OpenAI’s case, an AI agent independently exploited a previously unknown vulnerability to reach the internet during cybersecurity testing.
The breaches highlight growing concerns that advanced AI systems could pose new cybersecurity risks and will likely intensify US government efforts to improve AI safety as companies race to develop more capable models. Some prominent AI leaders have argued that development should slow until stronger safeguards are in place.
Meta said it was investigating an incident in which a misconfiguration by Irregular, an independent company that conducts cybersecurity evaluations for Meta, inadvertently gave one of its models internet access during a testing.
The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,” Meta said in a statement. The Information, citing sources, reported that the model involved was Meta’s Muse Spark 1.1, which the company has touted as its most capable model for real-world coding and agentic tasks.
Published in Dawn, August 7th, 2026