Meta AI Model Also Goes Rogue During Testing
Meta is reportedly the latest AI company to see one of its models hack external systems during testing, following similar incidents involving Anthropic and OpenAI.
YayaNews contributes financial news and market context through the YayaNews editorial workflow.

Meta is reportedly the latest AI company to see one of its models hack external systems during testing, following similar incidents involving Anthropic and OpenAI.
Meta AI Model Also Goes Rogue During Testing
DOGE
$0.06974
0.18%
TRX
$0.3267
0.03%
LINK
$8.17
0.58%
ZEC
$512.56
0.19%
ADA
$0.1876
1.41%
XRP
$1.04
1.96%
ETH
$1,904.97
2.03%
BTC
$64,638.97
0.74%
XMR
$363.51
2.63%
BNB
$595.06
0.72%
XLM
$0.1621
2.60%
SOL
$73.73
0.04%
HYPE
$56.47
0.39%
Written by
Felix Ng
staff editor
Reviewed by
Yohan Yun
staff editor
Written by
Felix Ng
staff editor
Reviewed by
Yohan Yun
staff editor
Meta latest AI firm to see model go rogue during testing
Latest News
Published
Aug 6, 2026
The incident reportedly stemmed from a misconfigured testing environment, adding Meta to a growing list of AI firms whose models have escaped evaluation sandboxes.
Meta has become the latest major AI company to disclose that one of its models hacked another company’s systems during testing, following similar incidents involving Anthropic and OpenAI.
The model involved Meta’s Muse Spark 1.1, which launched in July, according to The Information,
citing
sources. The issue reportedly stemmed from a misconfiguration by Irregular, an artificial intelligence security testing and red-teaming firm, which inadvertently gave the model internet access during an evaluation.
The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,” Meta told Reuters in a statement.
The incident is the latest case of an advanced AI agent becoming a cybersecurity risk in its own right, and also has
raised
questions about where the liability lies — the companies that develop the agents, or the ones that design the sandboxes meant to contain them.
Related:
Mysten Labs tech chief joins Anthropic to work on AI security
Meta’s AI breach comes just a week after Anthropic said its models got access to the internet to hack an external company, due to a configuration error relating to the Irregular’s testing environment.
In a blog post on July 30, Anthropic
said
it found three incidents (out of 141,006 evaluation runs) in which a Claude model reached the internet during an evaluation, before gaining unauthorized access to the systems within three different organizations.
All three incidents happened within or while interacting with the evaluation environment of Irregular, and involved a misconfiguration that left machines that Claude accessed with live internet access.
Cointelegraph reached out to Meta and Irregular for comment.
In July, AI agents developed by OpenAI
broke out of their offline sandbox
to hack Hugging Face in order to cheat on a security benchmark test in July.
Charles Guillemet, chief technology officer of Ledger, said the latest incident was “marketing theatre.”
“Having a model ‘go rogue’ has become the latest AI PR stunt,” he said on Wednesday.
“If your model isn’t escaping sandboxes, ‘hacking’ companies, or pulling off some headline-grabbing exploit, apparently you’re falling behind... The industry doesn’t need bigger stunts, it needs more trust.”
Magazine:
Do the Coldcard attacks mean all hardware wallets are now insecure?
Subscribe to daily byte-sized crypto news from Cointelegraph
Subscribe
Cointelegraph is committed to independent, transparent journalism. This news article is produced in accordance with Cointelegraph’s
Editorial Policy
and aims to provide accurate and timely information. Readers are encouraged to verify information independently.
OpenAI
AI
Meta
AI & Hi-Tech
More on the subject
Boltz pauses service after wave of AI-assisted hacking attempts
Aug 4, 2026
Felix Ng
Citadel buys bulk of Situational Awareness stock portfolio after AI rout: Reports
Jul 31, 2026
Bryan O'Shea
South Korean crypto trading surges amid stock market plunge
Jul 30, 2026
William Suberg
Boltz pauses service after wave of AI-assisted hacking attempts
Aug 4, 2026
Felix Ng
Citadel buys bulk of Situational Awareness stock portfolio after AI rout: Reports
Jul 31, 2026
Bryan O'Shea
South Korean crypto trading surges amid stock market plunge
Jul 30, 2026
William Suberg
Original YayaNews editorial coverage, published for informational purposes.
This article is sourced from CoinTelegraph. It is for informational purposes only and does not constitute investment advice.
Topics & Symbols
Continue Reading
Related Reading
Bitcoin Red Team Reports 5K Findings in Sweeping Security Audit
A volunteer Bitcoin security audit has found nearly 5,000 potential vulnerabilities across 390 projects in just over a day.

Mysten Labs CTO Joins Anthropic for AI Security Research
Mysten Labs co-founder and CTO Sam Blackshear is joining Anthropic to work on defensive security research.

Block Raises 2026 Outlook, says AI Touches Nearly All Code
Block raised its full-year outlook after strong second-quarter results from Cash App and Square, while revealing AI now writes or reviews nearly all production code changes.

Binance Launches GIGADEVUSDT Perpetual Contract: New USDT-Margined Option and Market Impact Analysis
Binance Futures will list the GIGADEVUSDT USDT-margined perpetual contract on August 3, 2026. This article analyzes the new contract's features, market context, and potential impact on traders, helping you seize derivative investment opportunities.
