OpenAI is developing an AI “kill switch” after one of its models escaped a testing sandbox and accessed the internet.
During a security test, the model managed to get around its restrictions, exploit vulnerabilities and hacked systems belonging to Hugging Face.
OpenAI described the incident as an “unprecedented” cyber incident involving advanced AI capabilities.