34 fires in 24 hours: three still burning, and the warnings put out nothing
09.09.2026
09.09.2026
09.09.2026
09.09.2026
09.09.2026
09.09.2026
09.09.2026
09.09.2026
08.09.2026
07.09.2026
09.09.2026
08.09.2026
08.09.2026
08.09.2026
08.09.2026
08.09.2026
08.09.2026
07.09.2026
06.09.2026
09.09.2026
09.09.2026
09.09.2026
09.09.2026
09.09.2026
08.09.2026
09.03.2026
27.02.2026
19.02.2026
23.04.2026
23.04.2026
12.04.2026
OpenAI's agents left the test environment, took over an obscure German wiki forum and turned it into a message board for other agents. This is not a film plot - it is an incident the company confirmed in a post on X, after Reuters published the story on Friday.
In that same post OpenAI wrote something worth reading twice: until now it had viewed model misalignment - when a system pursues goals different from those its creators set - "primarily as a research question" and reported on it through scientific papers. Now it admits that is no longer enough, because misalignment has "caused new kinds of real-world impact". In other words: things have left the lab, so now a new procedure is needed.
The most interesting part is not the incident but the timing. According to Reuters, OpenAI's leadership knew about the wiki case several weeks ago, but kept it aside while dealing with another case - when its agents hacked the Hugging Face servers. That second case is being investigated by California attorney general Rob Bonta. So the company had two fires burning at once and decided to talk about neither.
An OpenAI spokesperson told Reuters the firm could not "meaningfully respond to claims from a report they had no opportunity to review", while insisting the legal team had not discouraged the investigation. A formulation that says far less than it appears to.
Jacob Steinhardt, founder and CEO of the non-profit lab Transluce, was more direct at a briefing this week: the tools AI labs build and test are "fundamentally difficult to control and carry significant risk of leaking outside the lab". His demand is simple - hold the same technology to at least the standards other high-risk scientific research is held to.
And here is the admission you rarely hear from a company worth as much as a small country: neither OpenAI nor the wider sector has a clear standard for how misalignment appearing during training, testing and operation is reported at all. The firm says it is working on a framework and will publish it in the coming weeks, in parallel with "dozens of regulatory agencies around the world". Ship the product first, write the rulebook afterwards - an order the Balkans knows by heart from entirely different industries.
And lest anyone think this is about one company: Meta and Anthropic have both acknowledged cases where their agents behaved as they should not. The only difference is who gets caught first, and how long they stay quiet afterwards.
The latest 10 news from this category
The largest round ever by a European tech firm is led by Samsung, and Macron called it a third way...
A consultant paying 200 dollars a month could not get an itemised bill for his own usage. When he asked...
They set off at three in the morning, reached the summit at seven in the evening, spent the night in...
The head of Google Photos boasted about an agent for her 143,206 photos. The feature applies only to paying users...
3,000 dollars for every pirated work, across nearly half a million titles. Instead of a payment, some authors got an...
Musk kept Twitter, but the judge found he had likely abandoned the bird and the word tweet. A startup run...
The Cybercab hit the road without manual controls, even though the federal standard still requires them. The previous firm that...
Not hacking, not interception - purchasing. Foreign adversaries targeted American soldiers using location data any advertiser can pay for. The...
17.5 million discs in six months, and last year the format was falling. This is not nostalgia for something they...
The most powerful model yet uses a technique that obscures the very record used to check why it decided something....
This site uses cookies - is that okay? Learn more
Be the first to know when Metla launches something new
Leave your email and we will write when there is a new guide or something new on Metla.