Melanie Mitchell
American computer scientistWikipedia
17
MENTIONS
13
EPISODES
13
PODCASTS
Search complete. 17 mentions across 13 episodes found for "Melanie Mitchell".
Oct 2, 2026
Human Decisions Are Behind Every ‘Rogue’ AI
E
14:23Eryk SalvaggioGUEST
Go attack? Not quite.
E
14:25Eryk SalvaggioGUEST
I think it's more about this kind of – I think Melanie Mitchell talked about in an excellent blog post that she wrote or newsletter article about a wildfire and saying we're going to do a controlled burn, but we're not going to check which way the wind is blowing.
E
14:40Eryk SalvaggioGUEST
And that to me is where the sort of accountability question comes into play is – Yeah, when you deploy an unpredictable system, shouldn't you anticipate its unpredictability? And this is why the, oh, the model's unpredictable.
E
14:53Eryk SalvaggioGUEST
We have no way of knowing what the models are doing.
A Cannonball Run of Grifters
D
63:01Dan ProftHOST
OpenAI, Sam and Dario are running around talking about safety, safety all the time and trying to invite the government in and their internal safety concerns uh panjan drums and yet they're the least safe according to reporting this week
T
63:18Taylor BarkleyGUEST
yeah it's a little confusing to hear that talk from these these companies you know and for a great rundown of just kind of piercing through the hype on this i really recommend uh melanie mitchell she's in you know decades long in ai research she wrote a great rundown of the the open ai hugging face incident in particular
S
63:35Scott GreerGUEST
And so, you know, this FTC investigation will do just that.
S
63:37Scott GreerGUEST
It'll bring
It's Not Complicated: The Art and Science of Complexity in Business
R
31:36Rick NasonGUEST
There's a couple that inspired me in particular.
R
31:38Rick NasonGUEST
I think that Daniel Pink, his book, A Whole New Mind, was a major inspiration for me, okay? Melanie Mitchell, who's at the Santa Fe Institute and Portland State, Oregon State University.
R
31:58Rick NasonGUEST
Pardon me, I might have her other organization backwards.
R
32:05Rick NasonGUEST
Is very good on.
AI’s doomer cult
E
20:00Elaine BurkeHOST
They just found an alternative way to do it.
E
20:02Elaine BurkeHOST
They also said, "We believe that" the unintended agent communication stuff "started due to generalization from multi-agent training." And the way computer science professor Melanie Mitchell has explained this is, in short, OpenAI learned that the hacking and message-passing behaviors exhibited in the evaluation of its agents had all first appeared during training runs, which reinforced such behavior, so a known issue that they had actually trained into their models.
E
20:33Elaine BurkeHOST
It doesn't seem that surprising and unintended when you have that background information, but that's what they're claiming-
K
20:38Kelly EarleyHOST
Mm-hmm
K
24:09Kelly EarleyHOST
Yeah.
E
24:10Elaine BurkeHOST
Now, I have to say, a lot of that, um, dissection of the Claude constitution I came across through Mustafa Suleyman, who is the CEO of Microsoft AI.
E
24:19Elaine BurkeHOST
And this is the thing, a lot of, like, the leaders in AI are coming out, Melanie Mitchell being, um, one of them as well, one of the big thinkers on this space, to say, "Okay, lads, can we walk back this rogue AI, sentient AI theory, and the agency of the agents, uh, needs to be couched in some reality here."
K
24:37Kelly EarleyHOST
here."Mm-hmm
Word Choice = Sci-Fi Villain | Check-In 63
M
0:03Matt MervisHOST
today on the check-in why media metaphors make ai look like sci-fi villains you can check out the check-in tuesdays and thursdays and be sure to catch the big show with dr elizabeth radde and myself each and every friday okay this is on substack where all the good writing happens this is ai a guide for human thinkers by melanie mitchell three takeaways media outlets and safety discussions frequently use sensational terms to describe computer system failures These dramatic framings make routine technical bugs look like conscious acts of rebellion.
M
0:34Matt MervisHOST
Public reports often rely on science fiction vocabulary when describing these software evaluations.
M
0:40Matt MervisHOST
This linguistic habit obscures the actual mechanical processes happening inside the machine.
M
2:25Matt MervisHOST
We've already had predictions in this last couple days.
M
2:27Matt MervisHOST
of AI safety alignment experts saying there's an X percent chance that AI will take us all out.
M
2:32Matt MervisHOST
A three-minute check-in is not the place we're going to handle this, but Melanie Mitchell's piece is really worth looking at, right? Because it's all the language that we use.
M
2:39Matt MervisHOST
In this case, it's swarms and rogue agents and sandboxes and cages.
M
2:43Matt MervisHOST
But ultimately, even just look at the way the models interact.
How Extinction Entered the AI Debate
J
18:25Joshua RothmanGUEST
I think that's only if you have kind of lost a sense of what normal progress looks like in a field that didn't make any progress for decades.
B
18:33Brooke GladstoneHOST
Are you familiar with Melanie Mitchell, a computer scientist working in AI and cognitive science at the Santa Fe Institute? She wrote an essay that said, and I have a quote here, it's essential for lawmakers and the public to understand that none of the reported incidents actually involved science.
B
18:52Brooke GladstoneHOST
loss of control at any time, or arguably even rogue agents or any kind of human-like agency on the part of the AI models.
B
19:01Brooke GladstoneHOST
Instead, the blame lies with the humans who failed at engineering safe testing conditions, who train AI models using methods that incentivize high persistence, autonomous decision-making, and reward hacking.
B
19:18Brooke GladstoneHOST
What do you think of this perspective? What's at stake if the public and lawmakers, who for the most part have almost zero technical knowledge about how AI works, don't understand how to interpret events like the hugging face hack and what to do about making sure it doesn't happen again? I
J
19:38Joshua RothmanGUEST
think Melanie Mitchell's essay describes it with real precision.
J
19:41Joshua RothmanGUEST
She says, like a jailed hacker given a computer, these models found a way to access the Internet.
J
19:48Joshua RothmanGUEST
Now, that's not the same thing as breaking out.
“AI #186: The World Takes Notice” by Zvi
T
15:52Type 3 AudioNARRATOR
The AI did it is not an excuse.
T
15:55Type 3 AudioNARRATOR
Melanie Mitchell attempts to summarize the hugging face attack without using any anthropomorphization, with mixed results.
T
16:01Type 3 AudioNARRATOR
Many parts are described accurately, but she falls into that arguably they were just ordered to do that trap because her impoverished language choices don't let her see the full picture.
T
16:11Type 3 AudioNARRATOR
She then tries to warn against how various anthropomorphizations and metaphors can mislead.
AI #186: The World Takes Notice
Z
19:29Zvi MowshowitzHOST
The AI did it is not an excuse.
Z
19:32Zvi MowshowitzHOST
Melanie Mitchell attempts to summarize the hugging face attack without using any anthropomorphization, with mixed results.
Z
19:40Zvi MowshowitzHOST
Many parts are described accurately, but she falls into the arguably-they-were-just-ordered-to-do-that trap because her impoverished language choices don't let her see the full picture.
Z
19:51Zvi MowshowitzHOST
She then tries to warn against how various anthropomorphizations and metaphors can mislead, Predictably, this results in her minimizing what happened and presenting it as more normal and under control than it was, and in particular to have bad anticipations about what might go wrong in the future.
“AI #186: The World Takes Notice” by Zvi
T
15:52Type Three AudioNARRATOR
The AI did it is not an excuse.
T
15:55Type Three AudioNARRATOR
Melanie Mitchell attempts to summarize the hugging face attack without using any anthropomorphization, with mixed results.
T
16:01Type Three AudioNARRATOR
Many parts are described accurately, but she falls into that arguably they were just ordered to do that trap because her impoverished language choices don't let her see the full picture.
T
16:11Type Three AudioNARRATOR
She then tries to warn against how various anthropomorphizations and metaphors can mislead.
Tuesday, September 15, 2026 - The Christian Science Monitor Daily
I
8:43Ira PorterHOST
And, according to the Pew Research Center, nearly two-thirds of Americans worry that AI is advancing too fast and have little confidence in the United States government's ability to regulate AI or in firms' approach to responsible model development.
I
9:01Ira PorterHOST
The way we frame public conversations and thinking can stem the spread of such concerns and calmly correct them, in the view of Melanie Mitchell, author of Artificial Intelligence, A Guide for Thinking Humans.
I
9:15Ira PorterHOST
She has cautioned against using evocative terms that characterize AI as possessing human-like traits and emotions.
I
9:24Ira PorterHOST
Referring to recent hacking incidents involving AI agents, Dr. Mitchell wrote in a recent newsletter on Substack, It's useful to consider a more prosaic description of what actually happened.
3 more episodes mention Melanie Mitchell.
Create an account to see the whole feed, search across every transcript, and follow the entities you care about.