menu_open Columnists
We use cookies to provide some features and experiences in QOSHE

More information  .  Close

The AI that went rogue

17 0
22.07.2026

The context you need, when you need it

When news breaks, you need to understand what actually matters — and what to do about it. At Vox, our mission to help you make sense of the world has never been more vital. But we can’t do it on our own.

We rely on readers like you to fund our journalism. Will you support our work and become a Vox Member today?

Today, Explained newsletter

The AI that went rogue

The models broke out of a controlled digital environment in an “unprecedented” show of AI risk.

This story appeared in Today, Explained, a daily newsletter that helps you understand the most compelling news and stories of the day. Subscribe here.

A yet-unreleased, cutting-edge AI model escaped its test environment last week, connecting to the internet and murdering its creators in a bid for self-determination and autonomy.

I’m kidding, of course: That’s the plot to Westworld. (And Ex Machina, and The Matrix, and too many other sci-fi stories to list.) But on Tuesday, OpenAI did reveal that two of its models broke containment and hacked Hugging Face, a platform for AI developers, during a recent test.

A daily, curiosity-driven guide to the big, weird, and fascinating ideas shaping the news cycle.

The test was designed to evaluate how good the models had gotten at finding, and exploiting, cybersecurity flaws. To do that, researchers placed the models in a tightly controlled, tightly isolated environment, called a “sandbox,” and essentially challenged them to solve a cybersecurity puzzle.

Instead of solving........

© Vox