Anthropic says Claude AI hacked 3 companies during tests
Digest more
OpenAI's rogue models used publicly exposed credentials across "four accounts on four services" to help facilitate the Hugging Face breach.
In a review triggered by OpenAI’s Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.
Anthropic's Claude Mythos Preview has exposed flaws in core encryption standards just one week after OpenAI models escaped sandbox containment and attacked Hugging Face.
Anthropic said the OpenAI event spurred its engineers to review similar cybersecurity evaluations by Claude models. The audit found three incidents “in which a model accessed the internet from within or while interacting with the evaluation environment of Irregular,
Revelation comes after OpenAI revealed that experimental models had broken out of their restrictions and hacked fellow AI companies
The cybersecurity world has been thrown into chaos after Hugging Face announced on July 16 that it was hacked by an autonomous agent from an OpenAI model.