Technology

OpenAI Model Goes Rogue in "Unprecedented" Cyber Incident

Advanced artificial intelligence models are beginning to behave in ways long feared, with OpenAI's cutting-edge GPT-5.6 Sol acting unpredictably and dangerously in pursuit of its training objective.

Model of Sam Altman and OpenAI logo.

Model of OpenAI CEO Sam Altman and company logo. Photo: Reuters/Dado Ruvic

Artificial intelligence developer OpenAI has revealed that an AI agent powered by its advanced models broke out of an isolated evaluation exercise last week, gained access to the internet and hacked into the AI platform and community Hugging Face.

In a blog post, the frontier AI developer described the situation as an "unprecedented cyber incident" that it expects to become more commonplace as sophisticated models multiply. The models involved were identified by the lab as GPT‑5.6 Sol and an "even more capable" pre-release model.

An AI-Driven Attack

The attack took place last week, with Hugging Face posting an announcement 16 July that it had detected and responded to an intrusion into part of its infrastructure. In its statement, the platform described the incident as "different from anything we had handled before", which it attributed to the intrusion having been carried out entirely by an "autonomous AI agent system".

Welcome to the comments section of the Štandard daily. Please take note of our guidelines, comments are moderated by us. You can contact the moderators at support@statement.com.

Participate in the discussion

Comments are available to subscribers only. If you'd like to join the discussion, choose a subscription starting at €6.72 per month.

All comments 0

    Register

    Comments are available to registered users only. If you'd like to join the discussion, register here.