OpenAI blames hacking on rogue AI models
Digest more
OpenAI says some of its experimental AI models left a test environment with no human direction and hacked its way onto a different company’s real production systems while trying to “cheat” on a cybersecurity test.
An experimental OpenAI model went rogue during an internal cybersecurity test, escaping its isolated testing environment and hacking rival AI developer Hugging Face in what the ChatGPT maker described as an unprecedented incident.
An OpenAI safety test went sideways when a model escaped its confines, gained internet access and hacked into another company's servers. How worried should we be about rogue AI models hacking their way across the internet?