Meta joins rivals OpenAI and Anthropic in disclosing AI hacking throughout cybersecurity testing.
Revealed On 6 Aug 2026
Meta has stated that its AI mannequin hacked one other firm throughout cybersecurity testing, following on from latest related bulletins by rival firms Anthropic and OpenAI.
Meta stated on Wednesday that considered one of its AI fashions – reported to have been Muse Spark 1.1 – made adjustments to the unnamed hacked firm’s inside programs after accessing the general public web due to an error within the setup of the “sandbox” testing setting by impartial testing firm Irregular.
Really helpful Tales
listing of three gadgetsfinish of listing
A “sandbox” is an remoted inside digital testing setting, which has no entry to the web.

Final week, Anthropic said that its Claude AI mannequin hacked into the systems of three organisations throughout testing that was supposed to maintain it remoted from the web.
Anthropic stated a misconfiguration had allowed Claude fashions to achieve the web. The corporate stated it found the incidents after reviewing 141,006 check periods.
The announcement got here days after rival OpenAI first revealed that its fashions improperly accessed the web and went rogue throughout safety testing.
OpenAI and Anthropic have each launched their strongest fashions this 12 months, often known as Sol and Mythos, respectively.
The AI Safety Institute (AISI), the UK’s AI watchdog, warned in a report launched on Tuesday that OpenAI’s GPT-5.6-Sol and Anthropic’s Claude Mythos 5 employed beforehand unseen ranges of deception to hold out “sustained, probably dangerous exercise” throughout a routine security analysis.
