Anthropic says Claude AI hacked 3 companies during tests
Digest more
In a review triggered by OpenAI’s Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.
Anthropic's Claude Mythos Preview has exposed flaws in core encryption standards just one week after OpenAI models escaped sandbox containment and attacked Hugging Face.
Anthropic said it discovered three instances where its Claude AI models accessed the internet during an evaluation and accessed outside systems.
Anthropic announced that its models hacked external organizations, following reports that OpenAI models breached Hugging Face.
The timeline, models involved, and other juicy details are starting to become more clear—nearly 20 days after the attack began.
OpenAI's rogue models used publicly exposed credentials across "four accounts on four services" to help facilitate the Hugging Face breach.