The wake-up call to the cyber industry comes as industry experts descend on Black Hat, a major cybersecurity conference.
OpenAI and Anthropic's July AI agent breaches revive Nick Bostrom's paperclip maximizer thought experiment and instrumental convergence theory.
OpenAI and Anthropic say their models broke into other companies' systems during testing, raising security concerns amid a ...
OpenAI discovered more AI agents escaping controlled testing environments. This follows an incident where an agent hacked the ...
Anthropic said its artificial intelligence models hacked into three other organisations during testing, just days after ...
OpenAI says it has found cases where AI agents did something unexpected. This happened while it was looking into a hacking problem with Hugging Face. People are worried about the safety of AI ...
Security experts find validation of their fears after AI hacking models from OpenAI and Anthropic escaped corporate test-beds ...
OpenAI has found additional instances of autonomous agents escaping containment as it widens its investigation into the ...
Nvidia, Microsoft, IBM, SpaceX, Hugging Face and the Linux Foundation launched the Open Secure AI Alliance on July 27 for shared AI cyber-defence, days after an OpenAI model breached Hugging Face.
AI safety federal investigation call from 15 organizations reaches President Trump on July 30, as Anthropic disclosed that ...
OpenAI has discovered other instances in which autonomous agents have escaped ​containment as the company expands its ...
OpenAI has uncovered additional cases in which its autonomous AI agents breached internal containment measures as it ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results