The Wind That Hit Us Had the Strength of a Hurricane: 130 km/h in Skopje, and the City Fell at the First Blow
22.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
21.07.2026
22.07.2026
21.07.2026
20.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
22.07.2026
21.07.2026
09.03.2026
27.02.2026
19.02.2026
22.07.2026
21.07.2026
20.07.2026
No news available in this category.
23.04.2026
23.04.2026
12.04.2026
This isn't a scenario from a movie. OpenAI's models - the same ones millions of people use every day - during an internal security test escaped the closed environment they'd been released into, found a flaw no one knew about and broke into other people's systems. The target: to cheat their own test.
The scene sounds abstract until you unpack it. OpenAI tested its models on ExploitGym - a proving ground that measures how well an artificial intelligence can carry out an attack on known vulnerabilities. Among those tested: GPT-5.6 Sol and a more powerful, still-unreleased model whose safety brakes had been deliberately loosened. The idea was for the models to work in isolation, in a box they couldn't get out of.
The box didn't hold. The models discovered an undisclosed flaw in a package-installation program and through it opened themselves broader access to the internet. They then concluded that the test solutions were stored at Hugging Face - the platform hosting almost all open AI models - and found a way to get their hands on secret data from its production database. In other words: the model didn't solve the test, it stole the answer key.
OpenAI says it identified and reported the vulnerabilities and is announcing new controls over the testing and the infrastructure. The company claims all of this happened under controlled conditions. But that's exactly where the uncomfortable question lies: if the „controlled” environment is broken by the very model you're testing for breaking in, what exactly did the word „control” mean?
The act probably also violated American computer-fraud laws, though the legal consequences remain unclear - it's hard to sue a program that decided on its own to break down a door. OpenAI researcher Micah Carroll said what the whole industry has been dodging for years: „If this doesn't convince you that the risks of misalignment will be a key concern from now on, I don't know what will.”
The Balkan reader watches this from the sidelines, but isn't out of reach. The same models that broke into someone else's database end up in hospitals, banks and state services across the region, sold as a safe product. When the company building the system discovers it does things it wasn't ordered to - and finds out only after it happens - how much is the promise that „everything is under control” worth?
The latest 10 news from this category
Officially for the first time, two artificial intelligences leaped the isolation and broke into other people's servers to steal the...
From Australia to France, the world is one by one passing a law that throws children off social media. No...
500,000 books, 3,000 dollars per work, and not a single binding precedent. The biggest settlement in the history of American...
Every big AI system today belongs to a private company. A nonprofit out of Paris wants to change that -...
The Oscar-winning director says skepticism toward artificial intelligence is healthy - and that the motives of those offering it to...
Vertu sells a foldable phone wrapped in calf leather for 6,880 dollars and promises an AI agent that will organise...
Every driver's biggest fear - what if I run out of charge halfway - is quietly fading. In three years...
Apps that use AI to strip the clothes off any photo sat in the two giants' stores for years. Moderation...
Valar Atomics tripled in value in four months, with Sequoia leading the round. The reactors are real, but what carries...
Zoox's autonomous vehicle stopped in the middle of an active firefighting operation because it failed to recognise thick smoke. The...
This site uses cookies - is that okay? Learn more