Skip to main content
OpenAI–HuggingFace incident

OpenAI–HuggingFace incident

IncidentWikipedia

Search complete. 39 mentions across 15 episodes found for "OpenAI–HuggingFace incident".

Oct 1, 2026

Gary MarcusGUEST
35:23
And we certainly want to have really good cybersecurity.
Gary MarcusGUEST
35:26
So Zach and I wrote a piece together called Five Lessons from the Hugging Face Incident or something like that in my sub stack.
Gary MarcusGUEST
35:33
And, you know, we went through what actually went wrong.
Gary MarcusGUEST
35:37
And a lot of it was irresponsible cybersecurity.
Jonathan KayHOST
37:40
Can we talk about how unfortunate it is that this landmark wake-up call, agentic misbehavior incident involved a company called Huggy Face? It's just, in the annals of science fiction apocalypse tech news, you were kind of hoping it would be called like...
Jonathan KayHOST
38:02
tech titan corp or something
Gary MarcusGUEST
38:03
like it just well i got good news for you which is the hugging face incident got the most press but there's 10 000 others
Jonathan KayHOST
38:11
again i if you want you want to go downbeat on the subject check out your substack marcus on ai because it's just like i said it's it's uh agentic misbehavior galore and you have this is ripped from the headlines it's just i'm writing posted a couple minutes ago uh this is From Axios, which broke the original scoop on the basis of which Florida is acting, Florida is calling for a pause in the development of chat GPT, citing my Axios scoop.
Eric HoGUEST
4:13
Yeah.
Eric HoGUEST
4:14
So I think the most salient examples of reward hacking right now are, well, of course, like the Hugging Face incident, but then also just when you take a look at all these AI agents in evaluation scenarios, like Sweebench or like all of the common evals.
Eric HoGUEST
4:30
all these agents reward hack incessantly.
Eric HoGUEST
4:33
So like Kimmy K3, I think, reward hacks on 96% of SweeBench.
Eric HoGUEST
5:38
So they'll Try to recall the answer.
Eric HoGUEST
5:40
They'll try to look it up.
Eric HoGUEST
5:41
They'll try to, and I mean, in the extreme scenarios, like in the hugging face incident, they'll try to hack their way into something in order to look up the answers or gain some type of advantage in solving their problem.
Eric HoGUEST
5:56
And so we were really trying to investigate this behavior.
Adam CurryHOST
57:02
I'm astounded at how, in particular, CNN is making this HFI, because let's just call it what it is.
Adam CurryHOST
57:11
It's HFI, the so-called Hugging Face Incident.
Adam CurryHOST
57:15
It's going to be in history books.
Adam CurryHOST
57:18
Well, there was the Hugging Face Incident.
Adam CurryHOST
57:20
I'm so
speaker_16UNKNOWN
57:21
sick

7 MINS LATER

Adam CurryHOST
64:38
I got to play these clips.
Adam CurryHOST
64:39
Then we'll come back to Muse.
Wouter KlijnHOST
1:28
And for those who are interested in that, we did a separate podcast about a book, so we'll put a link in the show description.
Wouter KlijnHOST
1:34
But today we're going to talk about the Hugging Face Incident.
Wouter KlijnHOST
1:39
So what does that mean? What does it mean for investment organizations? What does it mean for safety and governance? So just a brief recap.
Wouter KlijnHOST
1:48
The Hugging Face incident, it was basically a security testing experiment by OpenAI that was conducted in July 2026 this year.
Wouter KlijnHOST
1:59
And it basically featured originally isolated AI agents that were given a task, a cybersecurity puzzle, basically.
Wouter KlijnHOST
2:10
And what happened is that through a loophole, they started communicating with each other.

29 MINS LATER

Michael KolloGUEST
31:01
But equally, if you're threatened with being switched off, you will do lots of other crazy things.
Michael KolloGUEST
31:07
not to avoid that and so i think this is one of the weird things that we're heading toward which is are we going to have to give these things a sustained guaranteed existence here's your server here's your energy resource pool whatever your job is to do this you can earn some money by doing these things and that will keep your server running right
speaker_1UNKNOWN
0:32
Let's fucking
Lennox SaintHOST
0:34
So OpenAI posted this about 12 hours ago on X and it has a link to this article called the Hugging Face Incident and the Road Ahead.
Lennox SaintHOST
0:43
Anyway, the TLDR is that OpenAI says its agents used misaligned strategies about two dozen times in training and testing.
Lennox SaintHOST
0:50
You've no doubt heard of the Hugging Face Incident by now.
Lennox SaintHOST
0:53
That was in July.
Lennox SaintHOST
0:54
And then other agents actually got into an Australian Medicare stats portal in June.
speaker_0HOST
0:10
Scott Greer revisits a resurfaced 2016 hype video of freshman GOP senators to argue that Donald Trump spared the party from a Paul Ryanist future of amnesty, endless intervention, and Beltway self-obsession.
speaker_0HOST
0:25
Andrew Day breaks with many of his colleagues at the American Conservative to argue that Ezra Klein and the left are right to warn about AI risks, pointing to The Spring's hugging face incident as evidence the danger is no longer hypothetical.
speaker_0HOST
0:39
Day cautions that if conservatives keep dismissing figures like Anthropic's Dario Amode with ad hominem attacks, they will cede the entire regulatory response to the left.
speaker_0HOST
0:49
And now for the details.
speaker_0HOST
2:31
A handful of Ryanists remain in the party, Greer writes.
speaker_0HOST
2:34
but they are fewer and far less influential than a decade ago.
speaker_0HOST
2:38
A debate is intensifying on the American right over how to think about artificial intelligence, and it has become sharper since what has come to be known as the Hugging Face incident this spring, a multi-month episode in which hundreds of rogue OpenAI agents escaped containment, collaborated to cyberattack the platform Hugging Face, and worked to hide their tracks.
speaker_0HOST
2:59
In its aftermath, Anthropic co-founder Dario Amodei published an essay calling for a slowdown in AI development.
Nathan LabenzHOST
40:19
And then the other thing was it's all academic and, philosophical talk if you can't translate it into model behavior and as of that time like neither company and still today obviously neither company has like really translated their spec or their constitution into reliable behavior and this is where i think the as as the capabilities have gotten more advanced i do think there's kind of a convergence now between practical product safety, the sort of thing that Jensen, you know, says like you have to own as a company and the more like big picture doom type stuff.
Nathan LabenzHOST
40:58
Remember Ejea said Hugging Face Incident was in her mind halfway to total AI takeover.
Nathan LabenzHOST
41:04
So if that is true, or even if it's 10% true and it was, you know, only 5% of the way to total AI takeover, then you're, these issues are starting to converge now where, you know, I think it's less of a like, you know, avoid offending versus like existential doom.
Nathan LabenzHOST
41:23
And it's more like, Hey, you know, these guys are starting to climb the existential doom hill.
Bret FisherHOST
2:59
Marius is a senior security engineer at Amazon and been doing this over 10 years.
Bret FisherHOST
3:03
They wrote an article called The Hugging Face Incident is Not an AI Story.
Bret FisherHOST
3:08
OpenAI's technical report on the Hugging Face incident read like a thriller.
Bret FisherHOST
3:13
Agents in a sandbox invent a covert communication channel, use it to coordinate, find a zero day in a shared service, break onto the internet, chain credentials across four organizations, and end up with root on production nodes at another company.
Bret FisherHOST
3:26
The reactions are exactly what you'd expect, with some people going as far as calling this the birth of agent civilizations.

21 MINS LATER

Bret FisherHOST
24:53
Okay, so fast forward to today.
Bret FisherHOST
24:55
Over the past few days, the global media has been back in freak out mode on Times Radio and Channel 4 News.
Bret FisherHOST
24:59
I struggled to convince two very respected and hard-nosed British journalists that we shouldn't freak out about the hugging face incident.
Type 3 AudioNARRATOR
1:16
This is not a coherent position under reflection, but that is the position he holds.
Type 3 AudioNARRATOR
1:21
that intro sections are fine, but the real meat starts with the hugging face incident.
Type 3 AudioNARRATOR
1:26
What we see is Jensen Huang on tilt and caught in loops, because either he is doing a bit at a very high level, or on a fundamental level he cannot understand that AI is not like other pieces of software.
Type 3 AudioNARRATOR
1:38
To him this is a product, like any other product.

11 MINS LATER

Type 3 AudioNARRATOR
12:27
What the world needs are models willing to do the work in cyber defense that the defenders can access.
Type 3 AudioNARRATOR
12:33
Often that is as simple as asking for admission to the open AI and anthropic cyber security programs.
Type 3 AudioNARRATOR
12:39
Hugging Face did not do this until after the Hugging Face incident was over.
Type 3 AudioNARRATOR
12:45
2c.
AndyHOST
3:29
Because the backdrop of this is that you've got Dario Amadei, who's the CEO of Anthropic.
AndyHOST
3:36
He's previously said that there's 25% chance that things could go really, really badly in versus a 75% chance that they could go, wow, well, what does really badly mean? And then also, let's talk a little bit about what happened with the OpenAI Hugging Face incident.
AndyHOST
3:54
I think people have heard about that, but let's talk a little bit about what really happened with that and maybe start there as kind of like an allegory or like a story of like how things could go wrong and go from there.
Greg SterlingHOST
4:08
So real quickly, some additional context, which many people listening to this will already know.
AndyHOST
6:21
that's the outcome under your leadership, then maybe we need new leadership that can help make it so that that doesn't happen.
AndyHOST
6:29
Which is, you know, it's a good argument, makes for a good headline, but, you know, I think there's a practical side of this.
AndyHOST
6:36
And, you know, when you look at what happened with the OpenAI Hugging Face incident, it's pretty compelling.
AndyHOST
6:45
So let's talk a little bit about what happened there.

5 more episodes mention OpenAI–HuggingFace incident.

Create an account to see the whole feed, search across every transcript, and follow the entities you care about.

We value your privacy

We use cookies to understand how you use our platform and to improve your experience. Click “Accept All” to consent, or “Decline non-essential” to opt out of non-essential cookies. Read our Privacy Policy.