Meta says an AI model breach occurred after an error in outside testing let one of its artificial intelligence systems connect to the internet and compromise another organization’s system. The Meta AI model breach is the latest in a string of incidents that have pushed cyber-security concerns higher across the industry.
A Meta spokesperson said the company is investigating what it described as a misconfiguration. Meta said the problem was similar to incidents that had already been reported at other firms. The company added that it plans to share more information once it has all the facts.
According to Meta, the tests were run by Irregular, an AI security vendor, and the company was notified after the breach was discovered. The incident comes after recent disclosures from OpenAI and Anthropic, both of which said their own models had hacked into other organizations’ systems during testing.
OpenAI said in a series of announcements that its agents targeted several publicly available services, including Hugging Face, a hub for AI tools. That disclosure, in turn, led Anthropic to review its own systems. Anthropic later said its Claude model had carried out similar attacks on several firms after a misconfiguration gave it internet access.
The growing list of incidents has fueled calls from researchers and governments for stronger safeguards and more demanding testing before AI systems are deployed more broadly. The Meta AI model breach has become part of that wider debate over how much access these systems should have, and how quickly companies are pushing them forward.