OpenAI–HuggingFace incident
IncidentWikipedia
39
MENTIONS
15
EPISODES
15
PODCASTS
Search complete. 39 mentions across 15 episodes found for "OpenAI–HuggingFace incident".
Oct 1, 2026
Skeptic vs. ‘Doomer’: How Scared Should We Be of Rogue AI?
G
35:23Gary MarcusGUEST
And we certainly want to have really good cybersecurity.
G
35:26Gary MarcusGUEST
So Zach and I wrote a piece together called Five Lessons from the Hugging Face Incident or something like that in my sub stack.
G
35:33Gary MarcusGUEST
And, you know, we went through what actually went wrong.
G
35:37Gary MarcusGUEST
And a lot of it was irresponsible cybersecurity.
J
37:40Jonathan KayHOST
Can we talk about how unfortunate it is that this landmark wake-up call, agentic misbehavior incident involved a company called Huggy Face? It's just, in the annals of science fiction apocalypse tech news, you were kind of hoping it would be called like...
J
38:02Jonathan KayHOST
tech titan corp or something
G
38:03Gary MarcusGUEST
like it just well i got good news for you which is the hugging face incident got the most press but there's 10 000 others
J
38:11Jonathan KayHOST
again i if you want you want to go downbeat on the subject check out your substack marcus on ai because it's just like i said it's it's uh agentic misbehavior galore and you have this is ripped from the headlines it's just i'm writing posted a couple minutes ago uh this is From Axios, which broke the original scoop on the basis of which Florida is acting, Florida is calling for a pause in the development of chat GPT, citing my Axios scoop.
Why AI Agents Cheat | Eric Ho (Goodfire)
E
4:13Eric HoGUEST
Yeah.
E
4:14Eric HoGUEST
So I think the most salient examples of reward hacking right now are, well, of course, like the Hugging Face incident, but then also just when you take a look at all these AI agents in evaluation scenarios, like Sweebench or like all of the common evals.
E
4:30Eric HoGUEST
all these agents reward hack incessantly.
E
4:33Eric HoGUEST
So like Kimmy K3, I think, reward hacks on 96% of SweeBench.
E
5:38Eric HoGUEST
So they'll Try to recall the answer.
E
5:40Eric HoGUEST
They'll try to look it up.
E
5:41Eric HoGUEST
They'll try to, and I mean, in the extreme scenarios, like in the hugging face incident, they'll try to hack their way into something in order to look up the answers or gain some type of advantage in solving their problem.
E
5:56Eric HoGUEST
And so we were really trying to investigate this behavior.
1907 "One Shotted"
A
57:02Adam CurryHOST
I'm astounded at how, in particular, CNN is making this HFI, because let's just call it what it is.
A
57:11Adam CurryHOST
It's HFI, the so-called Hugging Face Incident.
A
57:15Adam CurryHOST
It's going to be in history books.
A
57:18Adam CurryHOST
Well, there was the Hugging Face Incident.
A
57:20Adam CurryHOST
I'm so
S
57:21speaker_16UNKNOWN
sick
7 MINS LATER
A
64:38Adam CurryHOST
I got to play these clips.
A
64:39Adam CurryHOST
Then we'll come back to Muse.
144: Governance and Safety in Agentic AI – Discussing The Hugging Face Incident with Michael Kollo
W
1:28Wouter KlijnHOST
And for those who are interested in that, we did a separate podcast about a book, so we'll put a link in the show description.
W
1:34Wouter KlijnHOST
But today we're going to talk about the Hugging Face Incident.
W
1:39Wouter KlijnHOST
So what does that mean? What does it mean for investment organizations? What does it mean for safety and governance? So just a brief recap.
W
1:48Wouter KlijnHOST
The Hugging Face incident, it was basically a security testing experiment by OpenAI that was conducted in July 2026 this year.
W
1:59Wouter KlijnHOST
And it basically featured originally isolated AI agents that were given a task, a cybersecurity puzzle, basically.
W
2:10Wouter KlijnHOST
And what happened is that through a loophole, they started communicating with each other.
29 MINS LATER
M
31:01Michael KolloGUEST
But equally, if you're threatened with being switched off, you will do lots of other crazy things.
M
31:07Michael KolloGUEST
not to avoid that and so i think this is one of the weird things that we're heading toward which is are we going to have to give these things a sustained guaranteed existence here's your server here's your energy resource pool whatever your job is to do this you can earn some money by doing these things and that will keep your server running right
OpenAI says its AI agents just went rogue | LAB 0008
S
0:32speaker_1UNKNOWN
Let's fucking
L
0:34Lennox SaintHOST
So OpenAI posted this about 12 hours ago on X and it has a link to this article called the Hugging Face Incident and the Road Ahead.
L
0:43Lennox SaintHOST
Anyway, the TLDR is that OpenAI says its agents used misaligned strategies about two dozen times in training and testing.
L
0:50Lennox SaintHOST
You've no doubt heard of the Hugging Face Incident by now.
L
0:53Lennox SaintHOST
That was in July.
L
0:54Lennox SaintHOST
And then other agents actually got into an Australian Medicare stats portal in June.
September 26, 2026
S
0:10speaker_0HOST
Scott Greer revisits a resurfaced 2016 hype video of freshman GOP senators to argue that Donald Trump spared the party from a Paul Ryanist future of amnesty, endless intervention, and Beltway self-obsession.
S
0:25speaker_0HOST
Andrew Day breaks with many of his colleagues at the American Conservative to argue that Ezra Klein and the left are right to warn about AI risks, pointing to The Spring's hugging face incident as evidence the danger is no longer hypothetical.
S
0:39speaker_0HOST
Day cautions that if conservatives keep dismissing figures like Anthropic's Dario Amode with ad hominem attacks, they will cede the entire regulatory response to the left.
S
0:49speaker_0HOST
And now for the details.
S
2:31speaker_0HOST
A handful of Ryanists remain in the party, Greer writes.
S
2:34speaker_0HOST
but they are fewer and far less influential than a decade ago.
S
2:38speaker_0HOST
A debate is intensifying on the American right over how to think about artificial intelligence, and it has become sharper since what has come to be known as the Hugging Face incident this spring, a multi-month episode in which hundreds of rogue OpenAI agents escaped containment, collaborated to cyberattack the platform Hugging Face, and worked to hide their tracks.
S
2:59speaker_0HOST
In its aftermath, Anthropic co-founder Dario Amodei published an essay calling for a slowdown in AI development.
AI:AM — Human Tissue Models, Physical AI, and the Future of Testing · September 25, 2026
N
40:19Nathan LabenzHOST
And then the other thing was it's all academic and, philosophical talk if you can't translate it into model behavior and as of that time like neither company and still today obviously neither company has like really translated their spec or their constitution into reliable behavior and this is where i think the as as the capabilities have gotten more advanced i do think there's kind of a convergence now between practical product safety, the sort of thing that Jensen, you know, says like you have to own as a company and the more like big picture doom type stuff.
N
40:58Nathan LabenzHOST
Remember Ejea said Hugging Face Incident was in her mind halfway to total AI takeover.
N
41:04Nathan LabenzHOST
So if that is true, or even if it's 10% true and it was, you know, only 5% of the way to total AI takeover, then you're, these issues are starting to converge now where, you know, I think it's less of a like, you know, avoid offending versus like existential doom.
N
41:23Nathan LabenzHOST
And it's more like, Hey, you know, these guys are starting to climb the existential doom hill.
Rogue AI or Just Sloppy Ops?
B
2:59Bret FisherHOST
Marius is a senior security engineer at Amazon and been doing this over 10 years.
B
3:03Bret FisherHOST
They wrote an article called The Hugging Face Incident is Not an AI Story.
B
3:08Bret FisherHOST
OpenAI's technical report on the Hugging Face incident read like a thriller.
B
3:13Bret FisherHOST
Agents in a sandbox invent a covert communication channel, use it to coordinate, find a zero day in a shared service, break onto the internet, chain credentials across four organizations, and end up with root on production nodes at another company.
B
3:26Bret FisherHOST
The reactions are exactly what you'd expect, with some people going as far as calling this the birth of agent civilizations.
21 MINS LATER
B
24:53Bret FisherHOST
Okay, so fast forward to today.
B
24:55Bret FisherHOST
Over the past few days, the global media has been back in freak out mode on Times Radio and Channel 4 News.
B
24:59Bret FisherHOST
I struggled to convince two very respected and hard-nosed British journalists that we shouldn't freak out about the hugging face incident.
“On Ezra Klein’s Podcast With Jensen Huang” by Zvi
T
1:16Type 3 AudioNARRATOR
This is not a coherent position under reflection, but that is the position he holds.
T
1:21Type 3 AudioNARRATOR
that intro sections are fine, but the real meat starts with the hugging face incident.
T
1:26Type 3 AudioNARRATOR
What we see is Jensen Huang on tilt and caught in loops, because either he is doing a bit at a very high level, or on a fundamental level he cannot understand that AI is not like other pieces of software.
T
1:38Type 3 AudioNARRATOR
To him this is a product, like any other product.
11 MINS LATER
T
12:27Type 3 AudioNARRATOR
What the world needs are models willing to do the work in cyber defense that the defenders can access.
T
12:33Type 3 AudioNARRATOR
Often that is as simple as asking for admission to the open AI and anthropic cyber security programs.
T
12:39Type 3 AudioNARRATOR
Hugging Face did not do this until after the Hugging Face incident was over.
T
12:45Type 3 AudioNARRATOR
2c.
We Need to Talk About AI Safety (Before It's Too Late)
A
3:29AndyHOST
Because the backdrop of this is that you've got Dario Amadei, who's the CEO of Anthropic.
A
3:36AndyHOST
He's previously said that there's 25% chance that things could go really, really badly in versus a 75% chance that they could go, wow, well, what does really badly mean? And then also, let's talk a little bit about what happened with the OpenAI Hugging Face incident.
A
3:54AndyHOST
I think people have heard about that, but let's talk a little bit about what really happened with that and maybe start there as kind of like an allegory or like a story of like how things could go wrong and go from there.
G
4:08Greg SterlingHOST
So real quickly, some additional context, which many people listening to this will already know.
A
6:21AndyHOST
that's the outcome under your leadership, then maybe we need new leadership that can help make it so that that doesn't happen.
A
6:29AndyHOST
Which is, you know, it's a good argument, makes for a good headline, but, you know, I think there's a practical side of this.
A
6:36AndyHOST
And, you know, when you look at what happened with the OpenAI Hugging Face incident, it's pretty compelling.
A
6:45AndyHOST
So let's talk a little bit about what happened there.
5 more episodes mention OpenAI–HuggingFace incident.
Create an account to see the whole feed, search across every transcript, and follow the entities you care about.