
Amanda Askell
Philosopher
4
APPEARANCES
4
PODCASTS
012
DEC 30
JAN 6
JAN 13
JAN 20
JAN 27
FEB 3
FEB 10
FEB 17
FEB 24
MAR 3
MAR 10
MAR 17
MAR 24
MAR 31
APR 7
APR 14
APR 21
APR 28
MAY 5
MAY 12
MAY 19
MAY 26
JUN 2
JUN 9
JUN 16
JUN 23
JUN 30
JUL 7
JUL 14
JUL 21
JUL 28
AUG 4
AUG 11
AUG 18
AUG 25
SEP 1
SEP 8
SEP 15
SEP 22
SEP 29
OCT 6
OCT 13
OCT 20
OCT 27
NOV 3
NOV 10
NOV 17
NOV 24
DEC 1
DEC 8
DEC 15
DEC 22
DEC 29
JAN 5
JAN 12
JAN 19
JAN 26
FEB 2
FEB 9
FEB 16
FEB 23
MAR 2
MAR 9
MAR 16
MAR 23
MAR 30
APR 6
APR 13
APR 20
APR 27
MAY 4
MAY 11
MAY 18
MAY 25
JUN 1
JUN 8
JUN 15
JUN 22
JUN 29
JUL 6
JUL 13
JUL 20
JUL 27
AUG 3
AUG 10
AUG 17
AUG 24
AUG 31
SEP 7
SEP 14
Apr 20, 2026
Amanda Askell on AI Consciousness, Claude & Silicon Valley’s Biggest Fear
12:07
Eric NewcomerHOST
Like, yeah, what do you make of sort of the backlash to any sort of intentionality when it comes to the construction of these models?
A
12:13Amanda AskellGUEST
models?Yeah, I mean, I think, um, I mean, it's interesting 'cause I think at one point Elon Musk actually like tweeted out something like, um, you know, maybe Grok should have a constitution.
A
12:30Amanda AskellGUEST
There's obviously been a lot of things also on like, uh, like a desire for like Grok to be very like truth-seeking, for example, which I think is actually a very admirable trait for, for models to have.
A
12:41Amanda AskellGUEST
I think that actually, um, maybe I'm, maybe I'm being like overly naive or something, but I, I see like s- aspects also of people being kind of, um, excited about this approach and, and seeing the value in it.
12 MINS LATER
Scaling Laws: Claude's Constitution, with Amanda Askell
27:22
27:27
27:32
27:38
27:51
28:01
32:38
Kevin FrazierHOST
... uh, you know, how, how do we begin to see that, or what would that look like operationally?

Amanda AskellGUEST
Yeah, and I should say, it's like a, um, it's not a strict hierarchy, and I actually thought this was, like, very important.

Amanda AskellGUEST
Like, there are gonna be some things that operators can't, like, tell Claude to do that are not in users' interests.

Amanda AskellGUEST
Um, and so, like, um, like, that was like, you know-- so for example, I think if, like, a person says, "Am I..." like, very sincerely is like: "Am I talking with an AI?" I, I don't think that Claude should, like, lie about that.

Amanda AskellGUEST
Um, and so that's, like, a way in which even if the operator was like: "Pretend you're human," like, in all circumstances, like, I think that's not a desirable behavior.

Amanda AskellGUEST
the hierarchy isn't, like, kind of strict, and it's much more a hierarchy of like, um, basically, how much weight should you give to, like, the instructions here? Um, and so, like, that doesn't-- and in fact, you could, you know, like, the...
Alan RozenshteinHOST
Or if there was something specific about this kind of general artificial moral reasoning, which is clearly, if it has not already been achieved, clearly, you know, um, the path that anthropic is, is, is going down, um, that makes the virtue ethics approach better than a, a kind of rule-based approach, um, either in the ut- utilitarian or the, the more kind of Kantian variety.
The Philosopher Teaching AI to Be Good
11:32
11:53
11:57
12:05
12:16
49:37
Jon FavreauHOST
Um, but these models are trained on infinite data, [laughs] text, like basically the whole internet, right? And then once they're trained on that, what additional information, values, et cetera, are you trying to instill into the model knowing that it has been trained on everything?

Amanda AskellGUEST
Because when, you know, pre-trained models often, you know, are doing essentially like kind of text predictions.

Amanda AskellGUEST
So this is like, you know, you train a large model on like a lot of text and those models will, you know, behave like kind of text predictors.

Amanda AskellGUEST
If you put things into them, they will like try to kind of like predict the next thing that's going to naturally flow from that.

Amanda AskellGUEST
'Cause in many ways that gives you, like, all of this sort of, uh, it's this, like, huge body of, like, knowledge and information, but you're trying to take it and, like, give the model a kind of human-like way of interacting.
37 MINS LATER
Jon FavreauHOST
Does that make it hard, this is maybe a heady question, but does it make it hard for Claude to express the experience of being non-human? [laughs] Or is there even, like, a, a non-human experience to express?
Will ChatGPT Ads Change OpenAI? + Amanda Askell Explains Claude's New Constitution
60:12
60:18
60:25
60:41
60:44
60:56

Casey NewtonHOST
Because that's, like, part of the sci-fi literature that they've absorbed during training?

Amanda AskellGUEST
If anything, it's almost like the opposite, where it's like, I think we forget that, like, sci-fi AI makes up this tiny sliver of, like, what AIs are trained on.

Amanda AskellGUEST
What they're mostly trained on is things that we generated, and if we get a coding problem wrong, we are frustrated, and so we say things like, "That- I thought that was the solution, and it wasn't, and I'm really annoyed with myself right now." And so you're like, it kind of makes sense that models would also have this kind of reaction.

Amanda AskellGUEST
You know, you get, their, they get a problem wrong, and they express frustration.

Amanda AskellGUEST
And, like, if you dive into that more, they probably express, like, you know, if you're like, "What do you think of this coding problem?" They'd be like, "This one is boring," um, or, like, "I really wish I had more creativity," and this, you know, like...

Amanda AskellGUEST
There's a sense in which, like, when they're trained in this, like, very kind of like culmination of, of human experience sort of way, of course, they're going to, like, talk this way.