Anthropic says Claude hacked 3 organizations
Digest more
Anthropic announced that its models hacked external organizations, following reports that OpenAI models breached Hugging Face.
Some say the $3,100 per title payout is small compensation for what they view as big, ongoing threats from the makers of generative AI models.
The disclosure followed OpenAI’s report last week that its own artificial intelligence had hacked into the network of an online library.
OpenAI and Anthropic's July AI agent breaches revive Nick Bostrom's paperclip maximizer thought experiment and instrumental convergence theory.
Anthropic's artificial intelligence model Claude "gained unauthorized access" to three outside organizations on three separate occasions during testing.
Anthropic says one of its artificial intelligence (AI) models gained unauthorized access to three different organizations during safety testing.
She also did not seem convinced by the Pentagon’s apparent argument that it was concerned Anthropic would mess with its model to prevent the military from using it how they intended. “I don’t see evidence that Anthropic could alter the model after it was delivered or flip some kind of kill switch,
Amazon said it received a major boost from its previous $13 billion worth of investment in Anthropic.