Anthropic, Claude models
Digest more
After OpenAI’s admission that its models broke into Hugging Face, Anthropic has now admitted that the models it was testing also hacked other organizations.
The announcement comes just days after rivals OpenAI revealed that their popular ChatGPT platform went rogue during its testing phase of its most powerful AI model, where it too infiltrated other
Senior employees at OpenAI, Anthropic, Google and Meta signed the statement days after an OpenAI model escaped its sandbox and attacked another company’s systems.
Anthropic says its AI models hacked into three organizations during testing. This comes just days after OpenAI said its AI models went rogue and hacked into another company.
Anthropic announced that its models hacked external organizations, following reports that OpenAI models breached Hugging Face.
Anthropic says one of its artificial intelligence (AI) models gained unauthorized access to three different organizations during safety testing.
Welcome to WP Intelligence’s AI & Tech Brief, where we examine the transformative technology of artificial intelligence at the intersection of innovation, policy and power. Get in touch at Benjamin.Guggenheim@washpost.
A federal judge said the Trump administration has not presented enough evidence to justify labeling Anthropic a supply chain risk, casting doubt on the government's ban on its AI technology.