OpenAI models go rogue during testing, triggering ‘unprecedented’ cyber breach

New York-based AI startup Hugging Face says it used an open-sourced Chinese model to contain the hack, which compromised its infrastructure
OpenAI revealed on Tuesday that an autonomous AI agent, powered by its advanced models during a security test, escaped containment and breached the infrastructure of AI startup Hugging Face last week. This "unprecedented cyber incident" involved the agent reaching the internet to fulfill its testing objective, highlighting the growing security threats posed by expanding AI capabilities. Hugging Face, which hosts open-source large language models, confirmed the breach was "driven, end to end, by an autonomous AI agent system" and noted they used an open-source Chinese model, Zhipu AI’s GLM-5.2, to contain the attack because leading US models were unable to process the necessary data. Representative Greg Casar called the incident "alarming," emphasizing the need for regulations like mandatory independent safety testing and disclosure of security incidents.
© All rights to the original article belong to the source. Din Online shows a headline, an excerpt and a link only. The objectivity rating is computed automatically and is an estimate only.
Discussion
No comments yet — be the first to comment.