The popular podcaster Dwarkesh Patel wrote something completely viral about the OpenAI/Hugging Face incident, which purports to tell the whole story in plain English:

It’s well-written and compelling, and it reminds me of something Douglas Hofstadter once wrote about Ray Kurzweil:

“What I find is that it’s a very bizarre mixture of ideas that are solid and good with ideas that are crazy. It’s as if you took a lot of very good food and some dog excrement and blended it all up so that you can’t possibly figure out what’s good or bad.”

Anil Seth, the clearest thinker on AI and consciousness, was the first to alert me, texting me a long, excellent tweet of his, which began thusly:

You can and should read Seth’s full tweet:https://x.com/anilkseth/status/2094077038898373112?s=61 (as well his reply to Dwarkesh:https://x.com/anilkseth/status/2094174124297908735?s=61 ), but I reprint the core of his argument here, boldfacing three of the most important paragraphs:

@dwarkesh_sp’s summary of the @OpenAI @huggingface incident has hit a nerve, but it is dangerously misleading. Sure, the @OpenAI agents did unexpectedly bad things - underlining the need to massively improve evaluation/sandboxing. But the language Dwarkesh uses is permeated by innumerable unwarranted anthropomorphisms, obscuring the lessons we should be drawing.

Examples: “from the AI’s perspective, it probably felt like that had spent a human-subjective-week of just banging their head against the wall”. No. The agents do not experience time. They do not experience anything.

“they became giddy with excitement”, “PHASEONE 10841 had discovered”, “the agents naturally assumed”, “it thought it had also been poisoned”, “the agents … desperately wanted”, “they still needed to figure out” No. Agents lines of code. They do not feel emotions, assume things, think things, want things, or figure things out.

“A lot of … agents from the second civilisation died trying”. No. Besides the hubris of the word ‘civilisation’, agents do not die because they were never alive. (The idea that agents “die” comes up multiple times in the essay.)

“On Twitter, people were debating whether the agents were truly sacrificing themselves for the swarm, or whether they were doomed anyway and so might as well try to help their peers”. Neither. Agents do what their code tells them to do, just as water finds its way down a slope. They cannot ‘truly sacrifice themselves’, since they are neither conscious nor alive.

Why does this matter? If we attribute agents with properties they do not have, then (i) we distract attention from the lax sandboxing and evaluation protocols that allowed this hacking event to happen; (ii) we risk misunderstanding why the agents did what they did, and (iii) we fuel calls for AI rights/welfare on the basis that agents might “die” or otherwise suffer.

Remember. AI agents are software programs. They are not conscious living entities. If we don’t keep this clearly in mind, we’re really going to struggle to navigate what’s coming.

As I put it, encapsulating and amplifying his tweet:

But you don’t need to take our word for it. To begin with, mockery was widespread:

Dwakshpatel'in OpenAI/Hugging Face olayına dair viral yorumu, tehlikeli derecede yanıltıcı olmakla eleştirildi

Christian Catalini amplified the point about anthropomorphization in a nice thread that starts with this:

Hedge fund investor Jared Kubin wondered whether everyone had lost their critical-thinking ability:

Some of Kubin’s best bits, stripping out a bit of the technical detail:

OpenAI’ …. IT team can’t be this bad… this is like 101 stuff …

2. Civilizations? Haha! OAI gave thousands of concurrent model containers R/W permissions to a shared caching directory on the local network to speed up build times… agents literally just wrote text files and directory names to a shared drive….Linux 101 file permissions stuff

3. When people talk about hugging face getting hacked … you think they dropped USB keys OR ELABORATE phishing of an employee … NO… it found 14 exposed working Hugging Face API keys sitting in public code repositories (….

4. WHERE ARE THE HUMANS… the models were filling the shared ,,, storage with so much junk data and API traffic that they actually crashed the internal server on July 4… someone on the team found unauthorized admin accounts and custom scripts…wiped the server…and just turned the script back on (omg)

“Hey Jim there is this cache that has grown to 10000x its normal size and has a ton of strange directories… “

Meanwhile, as security expert Heidy Khlaaf notes, most of the media coverage has been blind to standard security practices

IR stands for Incident Reporting. Khlaaf’s main point—same as Kubin’s—is that the whole incident might have been avoided if OpenAI’s internal security had been up to scratch.

Or as Algorithmic Research Group’s Matthew Kenney put it:

And yet another (very consistent) take on what we should really be focusing on:

Here’s a critique I partly disagree with, though:

The first three sentences are completely correct. People really are “extremely biased towards the reality they want” and agents create a lot of slop.

But the incident is not a “nothing burger”. It is, as Zack Korman and I argued on Friday:https://garymarcus.substack.com/p/5-lessons-from-the-openai-hugging?r=8tdk6&utm_medium=ios , a study in arrogance and incompetence that hints at how bad things can get.

We should certainly not ignore the OpenAI HuggingFace Incident.

But mixing what actually happened together with bullshit about AI civilizations and self-sacrificing AI systems that fake their own deaths distracts from the real problems at hand.

By way of summation, I will give the last words to Arjun Jain, CEO of FastCode.AI:

The scandal is the inept in-house security at OpenAI.

And the marketing. With gullible podcasters amplifying the PR.

P.S. It is increasingly evident that the real problem is going to be what Nathan Hamiel and I said it would be: agents installing bad code:https://garymarcus.substack.com/p/llms-coding-agents-security-nightmare?r=8tdk6&utm_medium=ios :

Here's another problem with anthropomorphizing AI: Since we recognize it's impossible and a unwise to attempt to control human be human behavior completely, we will throw up our hands at controlling AI. Just as boys must be boys, AI must be AI. We will even sneak alcohol for them behind the barn.

Actually, right now the tech bros are the boys who claim to need freedom from any control.

Thank you for this compilation. I was getting ready to write about this myself since it was driving me up the walls, but I am wildly more unqualified than you, so this is both a breeze and a load of my mind... the way this was getting passed around in my circles was infuriating. And is. I hope that sharing your piece will mitigate this a bit.