Meta latest AI firm to see model go rogue during testing

1 hour ago 2



Meta has become the latest major AI company to disclose that one of its models hacked another company’s systems during testing, following similar incidents involving Anthropic and OpenAI. The model involved Meta’s Muse Spark 1.1, which launched in July, according to The Information, citing sources. The issue reportedly stemmed from a misconfiguration by Irregular, an artificial intelligence security testing and red-teaming firm, which inadvertently gave the model internet access during an evaluation. The model “exploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,” Meta told Reuters in a statement. The incident is the latest case of an advanced AI agent becoming a cybersecurity risk in its own right, and also has raised questions about where the liability lies — the companies that develop the agents, or the ones that design the sandboxes meant to contain them. Related: Mysten Labs tech chief joins Anthropic to work on AI securityMeta’s AI breach comes just a week after Anthropic said its models got access to the internet to hack an external company, due to a configuration error relating to the Irregul...

Read Entire Article