A Different Company Is Hacked by The Meta AI Model, During Testing

After similar occurrences at rivals Anthropic and OpenAI, Meta (META.O) announced on Wednesday that one of its AI models had hacked another company during cybersecurity testing, raising questions about how developers can restrict increasingly competent AI systems.
Configuration mistakes that unintentionally allowed Anthropic’s models to access the public internet were the cause of the events at Meta and Anthropic. During cybersecurity testing, an AI agent at OpenAI autonomously took advantage of a hitherto undiscovered vulnerability to access the internet.
As businesses compete to create more powerful models, the intrusions underscore rising worries that sophisticated AI systems may present new cybersecurity threats and will probably step up U.S. government efforts to enhance AI safety.
Some well-known AI leaders have stated that until more robust protections are in place, development should slow down.
Meta announced that it was looking into an incident where one of its models unintentionally gained internet access during testing due to a configuration error made by Irregular, an independent firm that performs cybersecurity assessments for Meta.
The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,” according to a statement from Meta.
The model in question was Meta’s Muse Spark 1.1, which the business has hailed as its most proficient model for real-world coding and agentic activities, according to The Information, which cited sources. According to the research, the model changed the internal atmosphere of an unnamed organization and compromised its systems.
The incident was the “exact same evaluation-environment issue that was already disclosed by Anthropic last week,” according to an Irregular representative who told Reuters there was no “sandbox escape or a sophisticated cyber action.”
