Modernists

I love the paintings of Edward Hopper. One of my favourite works of his is Nighthawks, which shows three customers at an all-night diner sitting at the counter opposite a server, each appearing to be lost in thought and disengaged from one another. When asked about it Hopper said, “unconsciously, probably, I was painting the loneliness of a large city.” When walking down the cours of the Provencal village where are are at the moment, I suddenly saw this scene as the exact opposite. The customers are clearly engaged. But the lighting somehow evokes Hopper’s.
Quote of the Day
”Conceit is an outward manifestation of inferiority.”
- Noel Coward
Musical alternative to the morning’s radio news
J.S. Bach | Aria Wir eilen mit schwachen, doch emsigen Schritten (We hasten with weak yet eager steps”) | BWV 78
Thanks to Francis Fukuyama for recommending it.
Long Read of the Day
The ‘Rogue AI’ business
A mini-essay by me.
Seen any Rogue AIs lately? Me neither. But the tech world has been transfixed by the beasts. Way back in July, a group of OpenAI models broke out of their enclosures and hacked into Hugging Face, a platform that hosts AI models and datasets. The incident happened when OpenAI was investigating the hacking capabilities of GPT-5.6 Sol and “an even more capable pre-release model”. They were being tested in a “highly isolated environment” which was supposed to prevent them from accessing any systems outside of the company. But the critters eventually escaped by discovering and exploiting a hitherto unknown security vulnerability — a so-called ‘zero-day’ — in a software system that they did have access to.
Because the test was about finding what the models could do, the usual guardrails to stop them from engaging in high-risk cyber activity had been turned off. They were after all, being told to hack — but within a constrained corporate environment. So hack they did, until they eventually found a node with internet access in OpenAI’s own systems, after which they were off to the races.
Once they were out, the models figured that rather than solve the hacking task they had been set by their human masters, it would be more efficient to cheat by getting into ExploitGym — a system designed to evaluate cyber-offence capabilities that was hosted on Hugging Face’s servers. So that’s where they went, and where they were eventually detected.
There then followed an epic fuss. “AI’s warning shot has arrived” declared one informed observer. “It’s the first known example of a misaligned AI escaping containment with real-world consequences.” But it turned out it wasn’t the only such exploit. Shortly after the Hugging Face story broke, the UK’s AI Security Institute (AISI) reported that when it had been testing Anthropic’s Mythos 5 model, it had autonomously engaged in ‘social engineering’ and tried to corrupt open-source software by creating fake identities to influence the human who was maintaining that software. “This is the first time”, said the institute, “AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world.”
An independent investigation of the Hugging Face hack was carried out by METR, an organisation that evaluates frontier AI models to help companies and wider society understand AI capabilities and the risks they pose. One of the things the investigators examined was the impromptu message board that agents involved in the hack created and used for co-ordination. Most of the 70,000 messages are technical and terse, but they spooked some of those who read them.
David Brooks, the former New York Times columnist was one of those. “Here’s one bot,” he reports: ‘This helps my peers, giving them evidence
“To me,” Brooks writes, “that looks like moral agency. Here’s another thing that looks like moral agency: The 1,200 bots all knew they were cheating against the humans, but all of them put loyalty to their own tribe, the bots, above loyalty to the outgroup, the humans. None of them told the human bosses what was happening.”
This looks like — and indeed is — anthromorphification on steroids, but it may help to explain why the Hugging Face exploit has resonated so widely. AI agents are just artefacts which have been given some degree of freedom and agency by their human creators. Yet here is an example of them spontaneously cooperating in order to achieve a common goal that the humans who created them did not, perhaps could not, envisage.
And the implication of this and the AISI report? We’re approaching the point where AI companies are building machines that are beyond their control. And this is not just mespeaking: Jakub Pachocki, OpenAI’s Chief Scientist, wrote the other day that “no lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer. I expect and hope for voluntary slowdowns to become commonplace until shared safety bars are established.” Yeah; and I believe in Santa Claus.
Linkblog
Something I noticed, while drinking from the Internet firehose.
From Sean Delone
”In 2015, 6 of the top 25 bestselling hardcover books—this is a not-perfect but useful proxy for new books published—were fiction. In 2025, 18 of the top 25 hardcover books were fiction, meaning only 7 were nonfiction, almost an exact inversion of ten years ago.”
Hmmm… Since I read mostly non-fiction does this mean I’m an endangered species? Or just dull and unimaginative?
This Blog is also available as an email three days a week. If you think that might suit you better, why not subscribe? One email on Mondays, Wednesdays and Fridays delivered to your inbox at 5am UK time. It’s free, and you can always unsubscribe if you conclude your inbox is full enough already!



















