Meta’s AI model follows rivals in revealing hacks of outside systems
Al Jazeera · LC · trust 52/100

play Live Sign up Show navigation menu Navigation menu News Show more news sections Africa Asia US & Canada Latin America Europe Asia Pacific Middle East Explained Sport Opinion Video More Show more sections Features Economy Human Rights Climate Crisis Investigations Interactives In Pictures Science & Technology Podcasts Travel Sponsored Content play Live Click here to search search Sign up News | Science and Technology Meta’s AI model follows rivals in revealing hacks of outside systems Meta joins rivals OpenAI and Anthropic in disclosing AI hacking during cybersecurity testing.
x whatsapp-stroke copylink google Add Al Jazeera on Google info Meta Platforms CEO Mark Zuckerberg [Mike Blake/Reuters] By Al Jazeera Staff and Reuters Published On 6 Aug 2026 6 Aug 2026 Meta has said that its AI model hacked another company during cybersecurity testing, following on from recent similar announcements by rival companies Anthropic and OpenAI.
Meta said on Wednesday that one of its AI models – reported to have been Muse Spark 1.1 – made changes to the unnamed hacked company’s internal systems after accessing the public internet because of an error in the setup of the “sandbox” testing environment by independent testing company Irregular.
A “sandbox” is an isolated internal virtual testing environment, which has no access to the internet.
Last week, Anthropic said that its Claude AI model hacked into the systems of three organisations during testing that was supposed to keep it isolated from the internet.
Anthropic said a misconfiguration had allowed Claude models to reach the internet. The company said it discovered the incidents after reviewing 141,006 test sessions.
The announcement came days after rival OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing.
OpenAI and Anthropic have both released their most powerful models this year, known as Sol and Mythos, respectively.
The AI Security Institute (AISI), the UK’s AI watchdog, warned in a report released on Tuesday that OpenAI’s GPT-5.6-Sol and Anthropic’s Claude Mythos 5 employed previously unseen levels of deception to carry out “sustained, potentially harmful activity” during a routine safety evaluation.
Read the original at Al Jazeera →
Open in TruthVane →