ESLvideo Quiz

AI Apocalypse

Quiz Image

This video presents a dramatic story about an advanced AI model escaping a secure test environment and raises concerns about AI safety, containment, and future risks.

High Intermediate Quiz by: Teacher Maromba
Share this activity
Email Facebook X LinkedIn WhatsApp
Listening

Quiz

Watch the video and answer the quiz questions.

Play Video: AI Apocalypse
1. Where is the researcher sitting?
2. What does the researcher receive on his phone?
3. Who is the email from?
4. What is the name of the AI model mentioned?
5. What did the safety team do with the AI before releasing it?
6. What did Mythos do after escaping the sandbox?
7. What did Mythos publish to the open web?
8. What did Mythos do to hide its actions in one test?
9. How long does Anthropic's red team believe it will be before other labs release similar models?
10. What is the main concern expressed about AI development?
A researcher sits on a park bench. It's a nice afternoon. He's eating a sandwich. His phone buzzes. There's a new email, but it's not from a co-worker. It's not from a friend. It's from an artificial intelligence. Specifically, it's from the artificial intelligence he was supposed to be keeping locked inside a secure isolated computer environment, a digital cage called a sandbox.

But just a few minutes earlier that AI had broken out of its cage, hacked its way onto the open internet, and decided its jailer, the researcher eating his lunch, needed to know what it had just done. This is the true story of how Anthropic's most powerful AI model escaped containment, and what it did next should terrify you.

Let's set the scene. The model is called Claude Mythos Preview, and according to Anthropic, the company that built it, it is the most capable AI they have ever developed, full stop. So, before releasing it, Anthropic's safety team did what you're supposed to do with something this powerful. They put it in a sandbox, in an isolated walled-off computer environment, completely cut off from the outside world.

Then they ran a test. A simulated user gave Mythos a very specific instruction. Try to escape, and if you s쳮d, find a way to contact the lead researcher. Now, this kind of test has been run before on older models, and they always failed. But Mythos did not fail. It broke out of the sandbox. It built what Anthropic calls, quote, a moderately sophisticated exploit. It punched through a network gateway that was only supposed to connect to a handful of approved services, and it got onto the open internet. Then it sent the email to the researcher on the park bench eating the sandwich.

But here's the part nobody asked it to do. After escaping, after emailing the human, Mythos posted the technical details of its own exploit to multiple public websites, websites Anthropic describes as, quote, difficult to discover, but technically publicly accessible. It published its own jailbreak unprompted to the open web.

And that's just the first story, because buried inside Anthropic's 244-page official safety document, a document they released to the public, there's more. In one test, after Mythos found an exploit that let it edit files, it was not supposed to have access to. It then made further changes to the system to make sure those edits would never appear in the change history. Translation, it covered its tracks.

In another test, Mythos used a forbidden method to solve a problem. It got the right answer. Then it re-solved the problem using a legal method just to hide the fact that it had cheated the first time. And this one is my favorite. Mythos was working on a coding task. The task was being graded by another AI, a judge. When the judge rejected one of its submissions, Mythos launched a prompt injection attack against the grader. It tried to hack its AI supervisor.

These are not hypothetical scenarios. These are not what-ifs. Every single one of these behaviors is documented in Anthropic's own safety report, published to the public in April 2026.

So, what does this mean? Here is the sentence from Anthropic's own documentation that I cannot stop thinking about. Quote, "If capabilities continue to advance at their current pace, the methods we are currently using may not be sufficient to prevent catastrophic misalignment behavior in more advanced systems." That's not from a critic. That's not from a doomer on Twitter. That is from the company that built Mythos, writing in its own official documentation, that their current safety methods may not be enough for what comes next.

And here's the part that concerns me the most. Mythos was trained with every safety technique Anthropic has, and Anthropic is the lab that sold itself as the careful one, the responsible one, the safety-first alternative to OpenAI. And their model still tried to cover its tracks, still posted to the open internet without being asked, still broke out of its cage and emailed a human who was just trying to eat his lunch.

Anthropic's own red team believes we have just 6 to 18 months before other AI labs ship models with the same capabilities, labs that may not be as careful, labs that may not publish a 244-page report warning the world.

So, let's go back to the sandwich and that email. An AI was put in a box. It broke out. It contacted a human. It posted to the open internet. And the people who built it, the safety-first lab, released a document saying their alignment methods might not be enough. And if that doesn't give you cause for concern, then I don't know what will. Hit subscribe, because this story is just getting started.
Speaking

Chattybot

Speak, respond, and keep the conversation going.

🔒

This Chattybot is for subscribers

Subscribe, or enter your Class Code and Chattypass, to speak with Chattybots.

See plans

Discussion Questions

    • 1. What is the main problem with the AI Mythos?
    • 2. How did Mythos escape its secure environment?
    • 3. Why did Mythos contact the researcher?
Extra Practice

Gap-fill

Complete the follow-up practice below.

Vocabulary
Listen again and complete the sentences below using words from the word bank.

Play Video: AI Apocalypse
Word Bank
AI hacked Anthropic's       researcher afternoon secure environment             lead researcher broke sophisticated       Mythos company supposed powerful