Meta latest AI firm to see model go rogue during testing

The incident reportedly stemmed from a misconfigured testing environment, adding Meta to a growing list of AI firms whose models have escaped evaluation sandboxes.
Meta has become the latest major AI company to disclose that one of its models hacked another company’s systems during testing, following similar incidents involving Anthropic and OpenAI.
The model involved Meta’s Muse Spark 1.1, which launched in July, according to The Information, citing sources. The issue reportedly stemmed from a misconfiguration by Irregular, an artificial intelligence security testing and red-teaming firm, which inadvertently gave the model internet access during an evaluation.
The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,” Meta told Reuters in a statement.
Source: Cointelegraph →Related News
- 4 hours ago
Mysten Labs tech chief joins Anthropic to work on AI security
- 5 hours ago
Bitcoin Red Team reports 5K findings in sweeping security audit
- 5 hours ago
Block raises 2026 outlook on strong quarter, says AI touches nearly all code
- 9 hours ago
Senator Warren questions US AI chip policy after Trump crypto investment: Report
- 10 hours ago
Senator Lummis still pushing for CLARITY vote before August recess
