Anthropic says Claude hacked 3 organizations
Digest more
Revelation comes after OpenAI revealed that experimental models had broken out of their restrictions and hacked fellow AI companies
Anthropic models during cybersecurity tests, Nvidia leading a global AI safety coalition, major funding deals, new AI model releases, and legal battles over AI ethics.
Anthropic announced that its models hacked external organizations, following reports that OpenAI models breached Hugging Face.
Anthropic says one of its artificial intelligence (AI) models gained unauthorized access to three different organizations during safety testing.
Anthropic said it discovered three instances where its Claude AI models accessed the internet during an evaluation and accessed outside systems.
Reports of problems are dropping sharply as service is restored.
Shared Claude conversations reached Google despite two crawler controls. Labs audit what their models say. Almost nobody audits what their products publish.