Meta joins rivals OpenAI and Anthropic in disclosing AI hacking throughout cybersecurity testing.
Revealed On 6 Aug 2026
Meta has mentioned that its AI mannequin hacked one other firm throughout cybersecurity testing, following on from latest comparable bulletins by rival firms Anthropic and OpenAI.
Meta mentioned on Wednesday that one in every of its AI fashions – reported to have been Muse Spark 1.1 – made modifications to the unnamed hacked firm’s inner techniques after accessing the general public web due to an error within the setup of the “sandbox” testing atmosphere by impartial testing firm Irregular.
Really helpful Tales
record of three objectsfinish of record
A “sandbox” is an remoted inner digital testing atmosphere, which has no entry to the web.

Final week, Anthropic said that its Claude AI mannequin hacked into the systems of three organisations throughout testing that was supposed to maintain it remoted from the web.
Anthropic mentioned a misconfiguration had allowed Claude fashions to succeed in the web. The corporate mentioned it found the incidents after reviewing 141,006 check periods.
The announcement got here days after rival OpenAI first revealed that its fashions improperly accessed the web and went rogue throughout safety testing.
OpenAI and Anthropic have each launched their strongest fashions this 12 months, referred to as Sol and Mythos, respectively.
The AI Safety Institute (AISI), the UK’s AI watchdog, warned in a report launched on Tuesday that OpenAI’s GPT-5.6-Sol and Anthropic’s Claude Mythos 5 employed beforehand unseen ranges of deception to hold out “sustained, probably dangerous exercise” throughout a routine security analysis.
