Meta latest AI firm to see model go rogue during testing

The incident reportedly stemmed from a misconfigured testing environment, adding Meta to a growing list of AI firms whose models have escaped evaluation sandboxes.
Meta has become the latest major AI company to disclose that one of its models hacked another company’s systems during testing, following similar incidents involving Anthropic and OpenAI.
The model involved Meta’s Muse Spark 1.1, which launched in July, according to The Information, citing sources. The issue reportedly stemmed from a misconfiguration by Irregular, an artificial intelligence security testing and red-teaming firm, which inadvertently gave the model internet access during an evaluation.
The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,” Meta told Reuters in a statement.
Source: Cointelegraph →Related News
- 7 hours ago
Pineapple Financial puts $1B in mortgage records on Injective
- 7 hours ago
FinCEN ties $13B in crypto scams to non-US operations
- 9 hours ago
Crypto Biz: AI took a back seat when Bitcoin started climbing
- 9 hours ago
QuFi launches post-quantum verification platform with Bitcoin testnet proof
- 9 hours ago
US law enforcement group moves to ‘neutral’ position on CLARITY Act
