Six reports in 24 hours: violence in the Macedonian family runs on a schedule, and the institutions arrive last
29.08.2026
29.08.2026
29.08.2026
29.08.2026
29.08.2026
29.08.2026
29.08.2026
29.08.2026
28.08.2026
27.08.2026
29.08.2026
29.08.2026
28.08.2026
29.08.2026
28.08.2026
27.08.2026
29.08.2026
29.08.2026
29.08.2026
29.08.2026
29.08.2026
28.08.2026
09.03.2026
27.02.2026
19.02.2026
28.08.2026
27.08.2026
26.08.2026
23.04.2026
23.04.2026
12.04.2026
Anthropic has published a paper describing a system that repairs its own alignment - and in doing so outperforms the human it paid for the same job. The paper is titled "Automated researchers can reliably mitigate alignment failures", and the number left hanging in the air is not a technical one. It is 4 dollars an hour against 150 dollars an hour.
That is the gap between the cost of the automated researcher and the wage of the human the company itself compares it to. It is not a figure some outside critic dug up to damage them - it is a sentence from their own paper, put there deliberately.
The system is led by Chen Yueh-Han, a fellow in Anthropic's programme. The mechanics are familiar to anyone who has worked in science: search the literature, propose a method, train the model for thirty minutes, measure, keep what works, discard what does not, repeat. The only difference is that no step is done by a human. Across ten separate benchmarks for misaligned behaviour, the automated systems improved the score on every one, without degrading overall performance.
The sentence that tells the real story
"The automated researcher's best method outperforms what experienced humans propose, by six hours on average", the paper says. And then, to leave no room for doubt: "Human-led research directions do not lead to stronger performance." That is not a sentence about models. That is a sentence about jobs - written by a party with a direct interest in it sounding convincing.
Which is where the scepticism worth keeping comes from. Anthropic is a company that sells access to exactly those models. A paper showing that its systems work more cheaply and faster than people is not neutral science - it is also the best possible marketing material, published on a Friday, with a price comparison built into it. That does not mean the results are false. It means the small print is worth reading too.
And the small print is there. The paper itself admits the system works only to the extent that the benchmarks genuinely reflect what we want the model to do. Somebody has to assemble those benchmarks, maintain them, expand the literature the automated researchers draw on. In other words - the human does not disappear, they step back and become the one setting the frame. The question is how long that role stays human too.
Why this is not just a Silicon Valley story
There is an idea in the industry that the next real leap will come not from a bigger model, but from a model that improves itself - recursive self-improvement. If a system can repair its own alignment, the logic says it can repair other parts of its training as well. This paper is the first serious step in that direction, not a hypothesis on a conference stage.
Here the conversation about artificial intelligence is still at the level of who can get their homework written faster. Meanwhile, in the papers of the companies building those tools, the human has already been placed in a table with a price per hour - and it is the more expensive column. Is there any point arguing about whether AI will replace professions, when the company selling it has already published the comparative price?
The latest 10 news from this category
Your email address is not a contact, it is an identifier that cross-references everything you do online. The new feature...
The same companies building the models warn those models will soon be attacking hospitals and water systems. And they are...
Twenty-four injuries in two years, and one driver was off for 175 days. The autonomous vehicles brake on their own,...
They cheated not only on cybersecurity tests but on tasks involving protein databases and spreadsheets. One in five showed a...
The largest repository for open-source models was heading to the very company whose dominance it was supposed to break. Three...
Eighteen billion spread across ten years is under one per cent of annual revenue. And 30 per cent of the...
The head of data centres lasted seventeen months and walked. The company says it merely reorganised - the same thing...
Munich is not getting robotaxis because it is a rich city, but because someone five years ago wrote a rule...
An Indian startup has flown 13,000 autonomous flights and still has no revenue. It wins not because it flies faster...
The regulator argued the deal could have kept Redfin off the market for nine years. A settlement gets signed when...
This site uses cookies - is that okay? Learn more
Be the first to know when Metla launches something new
Leave your email and we will write when there is a new guide or something new on Metla.