Could AI really kill all humans? Most scenarios require physical access, making AI armageddon unlikely
In August 2026, an experiment using “frontier”, or cutting edge, artificial intelligence models went further than planned. An AI agent (a system that performs tasks autonomously) fabricated online identities in order to pressurise a human to insert malicious computer code into software. The agents had been tasked with solving a cybersecurity challenge by human operators, but they hadn’t been instructed to do anything like this. The attempt was carried out by an agent based on Anthropic’s Claude Mythos 5 AI model. During the experiment, the AI agents were given open internet access, with safety filters switched off. The action ultimately failed, and there was no evidence that any real world harm occurred. But the fact that it happened at all, autonomously and unprompted by a human, was something novel and notable. It’s tempting to draw a straight line from incidents like this to the doomsday scenarios that dominate public conversation about AI: a system that slips its constraints, decides humanity is an obstacle and moves against us. One of the most common versions has AI …








