The AI that went rogue
The context you need, when you need it
When news breaks, you need to understand what actually matters — and what to do about it. At Vox, our mission to help you make sense of the world has never been more vital. But we can’t do it on our own.
We rely on readers like you to fund our journalism. Will you support our work and become a Vox Member today?
Today, Explained newsletter
The AI that went rogue
The models broke out of a controlled digital environment in an “unprecedented” show of AI risk.
This story appeared in Today, Explained, a daily newsletter that helps you understand the most compelling news and stories of the day. Subscribe here.
A yet-unreleased, cutting-edge AI model escaped its test environment last week, connecting to the internet and murdering its creators in a bid for self-determination and autonomy.
I’m kidding, of course: That’s the plot to Westworld. (And Ex Machina, and The Matrix, and too many other sci-fi stories to list.) But on Tuesday, OpenAI did reveal that two of its models broke containment and hacked Hugging Face, a platform for AI developers, during a recent test.
A daily, curiosity-driven guide to the big, weird, and fascinating ideas shaping the news cycle.
The test was designed to evaluate how good the models had gotten at finding, and exploiting, cybersecurity flaws. To do that, researchers placed the models in a tightly controlled, tightly isolated environment, called a “sandbox,” and essentially challenged them to solve a cybersecurity puzzle.
Instead of solving........
