It is one of the first publicly disclosed cyber-attacks carried out by AI without direct human involvement.

OpenAI is best known for its chatbot ChatGPT, which is used by hundreds of millions of people every week OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test. The ChatGPT-maker said its agent - an AI system which can operate alone after some human instruction – was being tested in a controlled environment, but found vulnerabilities and managed to escape. They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems. OpenAI said the incident was "unprecedented", external, and it was conducting an investigation alongside Hugging Face, whose boss Clement Delangue said in a post on X it was "mind-blowing that all of this happened autonomously". "The investigation is ongoing, and we'll share more learnings from what might be the first incident of its kind," Delangue added. Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, told BBC Radio 4's Today programme that the security tests - called sandboxes - are "supposed to be secure environments where you can see what the models are capable of". "In this case, it looks like OpenAI didn't make a secure enough sandbox," she added.