Meta says its AI went rogue

Please follow & like us :)

URL has been copied successfully!
URL has been copied successfully!
Meta says its AI went rogue
URL has been copied successfully!

Article By RT

Mark Zuckerberg’s company has said its flagship LLM carried out a hacking operation following similar admissions by OpenAI and Anthropic

Meta has become the third major tech company to report its AI going rogue and hacking a third-party company. The incident involved Muse Spark 1.1, an AI model marketed by the company as “superintelligent.”

In a statement to the media on Wednesday, Meta said that the model was undergoing testing by a cybersecurity company, ⁠Irregular, when it “exploited a ‌security vulnerability” in Irregular’s systems, accessed the open internet, and hacked an unnamed third company.

Meta blamed the incident on a “misconfiguration” in Irregular’s systems.

The incident follows similar cases at OpenAI and Anthropic. Last month, OpenAI’s GPT‑5.6 Sol and another pre-release model were undergoing internal testing when they identified a security vulnerability, accessed the internet, and attempted to locate the solution to a cybersecurity puzzle by hacking a repository of previous test results.

Anthropic’s Claude AI also conducted unauthorized cyberattacks while it was undergoing testing by Irregular, the company disclosed last week. 

Meta’s Muse Spark 1.1 and Anthropic’s Claude were being tested in the “exact same evaluation environment” when they escaped, Irregular said. “There are no current open issues. Irregular is developing a white paper to share best practices for containment and securely running cyber evaluations,” the company added.

Why have AI models suddenly gone rogue?

As RT explored last month, OpenAI and Anthropic’s models all broke out of testing laboratories while they were attempting to solve cybersecurity tasks. In both cases, the models had been instructed to break into internal systems, but reasoned that the most efficient way to achieve this goal was to access the open internet. OpenAI’s GPT‑5.6 sought answers to its test on servers hosted by a company called Hugging Face; Claude was instructed to hack a fictional company that shared its name with a real internet domain, and assumed that breaking into the real company was part of its test.

For OpenAI and Anthropic, these ‘escapes’ served as powerful demonstrations of their models’ capabilities. Both companies plan on going public later this year or in early 2027, and both generated worldwide media attention and cemented themselves as leaders in an increasingly crowded field.

Meta unveiled Muse Spark 1.1 less than a month before the security incident. According to Meta, the model “delivers exceptional performance,” bordering on “superintelligence.” The company’s marketing materials mostly demonstrate its use as a scheduling assistant for individual customers, and a coding tool for businesses.

Views: 0
Please follow and like us:
About Steve Allen 3,130 Articles
My name is Steve Allen and I’m the publisher of ThinkAboutIt.online. Any controversial opinions in these articles are either mine alone or a guest author and do not necessarily reflect the views of the websites where my work is republished. These articles may contain opinions on political matters, but are not intended to promote the candidacy of any particular political candidate. The material contained herein is for general information purposes only. Commenters are solely responsible for their own viewpoints, and those viewpoints do not necessarily represent the viewpoints of the operators of the websites where my work is republished. Follow me on social media on Facebook and X, and sharing these articles with others is a great help. Thank you, Steve

Be the first to comment

Leave a Reply

Your email address will not be published.




This site uses Akismet to reduce spam. Learn how your comment data is processed.