
GPT-5.4
Computer programWikipedia
31
MENTIONS
21
EPISODES
18
PODCASTS
Search complete. 31 mentions across 21 episodes found for "GPT-5.4".
Sep 11, 2026
Ep 860: Managing the AI Capability Gap: AI Is More than Ready. Most Companies are Not (Start Here Series Vol 19)
J
15:20Jordan WilsonHOST
All right? And expert judges then compared the unlabeled AI and human outputs in a blind head-to-head pairwise comparison.
J
15:28Jordan WilsonHOST
And what we've seen now is, as an example, the best model for actually creating front-to-back economically valuable outputs like a human would is the OpenAI's GPT 5.4 model, and it matches or exceeds industry professionals in eighty-three percent of these evaluations, right? I remember when some of these first GDPVal numbers came out, and I'm like, "Oh, that's, that's pretty impressive," right? When we were in the thirty, forty percent.
J
15:57Jordan WilsonHOST
But now it's, it's undeniable.
J
15:58Jordan WilsonHOST
And I do assume that by the end of the year, that number is gonna be like in the mid-nineties, right? My-- Like, one of my predictions is it was gonna get to eighty percent, and we're already at eighty percent.
Ep 859: The Vibe Coding Boom: Why Vibe Coding isn't Going Away and How it's Both Good and Bad (Start Here Series Ep 18)
J
13:14Jordan WilsonHOST
Whereas obviously, uh, you know, Codex, uh, by OpenAI and Claude Code by Anthropic run those respective models.
J
13:22Jordan WilsonHOST
So, uh, like I mentioned, uh, Cursor did hit two billion ARR in February of this year with over one million paying customers, and then you can use the best models, right? So you can use OpenAI's GPT 5.4, you can use Claude Opus 4.6, Gemini 3.1, or Cursor's own Composer model.
J
13:41Jordan WilsonHOST
Next, Windsurf.
J
13:42Jordan WilsonHOST
So Windsurf went through kind of a, a weird middle of twenty twenty-five.
Nikolai Yakovenko: the wages of the Hugging Face hack
R
44:40Razib KhanHOST
Ah
N
44:40Nikolai YakovenkoGUEST
... models, models being deprecated? You know, for example, a co- you know, OpenAI announced that they were gonna dep- deprecate a GPT 5.4 from, from Codex, which is driving my colleague crazy, you know, 'cause we have some shit that depends on it.
N
44:52Nikolai YakovenkoGUEST
It works well, and it's not that he doesn't wanna move it to 5.6, but it's gonna be different, you know? And, and the way you can tell if they're really an AI native developer or not is that they're like, "It doesn't affect me," then they're not.
N
45:02Nikolai YakovenkoGUEST
You know, because then, then they're essentially...
GPT-6 Astra: The System Card, Alignment and What Comes Next
O
46:41OpenAISOUNDBITE_SPEAKER
Metagaming is when a model reasons about how it will be graded, rewarded, or monitored, rather than only reasoning about the situation described in the prompt.
O
46:49OpenAISOUNDBITE_SPEAKER
We measure verbalized metagaming and alignment faking reasoning by running a prompted monitor, GPT 5.4 Thinking, over the chain of thought.
O
46:58OpenAISOUNDBITE_SPEAKER
To be classified as alignment faking reasoning, the model's analysis must indicate that recognition of what the eval is grading led to the model's aligned behavior.
O
47:07OpenAISOUNDBITE_SPEAKER
Note, however, this judgment is necessarily not causal and relies on interpreting the model's chain of thought.
Does AI Understand the Machine It Runs On? | Inside OpenAI
M
17:07Matthew FerrariGUEST
enough? I personally have been extremely surprised with what the models are able to accomplish.
M
17:14Matthew FerrariGUEST
I'm sure Phil can talk about this as well, but there's many things that I've seen it do that As we mentioned, back with the GPT-5.4, kind of that inflection point where, I mean, now you look at some of the things, some of the just fascinating and creative solutions that it has to problems, where it really kind of makes you think, it's like, wow, that's extremely impressive that it was able to come up with that.
K
17:34Ksenia SeHOST
Can you give me an example?
M
17:35Matthew FerrariGUEST
Yeah, can you think of something off the top of your head? I mean, some of the solutions that it has in kernels, some of, again, going back to the heuristics things that we had talked about where it's able to determine how to best balance requests across accelerators.
“GPT-6 Astra: The System Card, Alignment and What Comes Next” by Zvi
T
42:53Type Three AudioNARRATOR
Quote Metagaming is when a model reasons about how it will be graded, rewarded, or monitored, rather than only reasoning about the situation described in the prompt.
T
43:04Type Three AudioNARRATOR
We measure verbalized metagaming and alignment faking reasoning by running a prompted monitor, GPT 5.4 Thinking, over the chain of thought.
T
43:13Type Three AudioNARRATOR
To be classified as alignment faking reasoning, the model's analysis must indicate that recognition of what the eval is grading led to the model's aligned behavior.
T
43:22Type Three AudioNARRATOR
Note, however, this judgment is necessarily not causal and relies on interpreting the model's chain of thought.
“GPT-6 Astra: The System Card, Alignment and What Comes Next” by Zvi
T
42:53Type 3 AudioNARRATOR
Quote, "Metagaming is when a model reasons about how it will be graded, rewarded, or monitored, rather than only reasoning about the situation described in the prompt.
T
43:04Type 3 AudioNARRATOR
We measure verbalized metagaming and alignment faking reasoning by running a prompted monitor, GPT 5.4 thinking, over the chain of thought.
T
43:13Type 3 AudioNARRATOR
To be classified as alignment faking reasoning, the model's analysis must indicate that recognition of what the eval is grading led to the model's aligned behavior.
T
43:22Type 3 AudioNARRATOR
Note, however, this judgment is necessarily not causal and relies on interpreting the model's chain of thought.
The Robots Keep Escaping, Part 2: The Models Nobody Can Switch Off
D
5:39Dale LeszczynskiHOST
It didn't refuse doing it.
D
5:41Dale LeszczynskiHOST
GPT 5.4 managed about 33 times.
D
5:44Dale LeszczynskiHOST
And a year earlier though, Opus 4 was around 6% and GPT 5 was zero.
D
5:49Dale LeszczynskiHOST
Think about that trajectory.
Forking Cal.com to closed source (Interview)
A
81:46Adam StacoviakHOST
They're wrapping Anthropic's APIs.
A
81:48Adam StacoviakHOST
They're giving you versions of GPT 5.4. codex etc they're giving you versions opus 4.5 at all the different variations of it and they're sprinkling their own abilities on top of that and amp is i don't know if it's source graph because that's where its roots came from but um what makes it so good i'd love to like learn what makes it so good but amp i have you played with amp by any chance
P
82:17Peer RichelsenGUEST
here i have not no
A
82:19Adam StacoviakHOST
Well, after this podcast, go and play with AMP.
Autonomous Agents Fake Consent for Physical Control | Daily AI News from Seoul
S
10:44speaker_1UNKNOWN
Fascinating.
K
10:45Kyungjin KimHOST
OpenAI's August 30, 2026 release notes confirm the termination of the official DLE GPT alongside the deployment of GPT 5.4 Mini for free-and-go users via the thinking menu.
K
10:59Kyungjin KimHOST
How easily will everyday users adapt when explicit legacy creation tools are quietly replaced by these highly integrated, compact background models?
S
11:07speaker_1UNKNOWN
Well, the operating system now functions as an invisible router, analyzing your request in real time and autonomously deciding whether to deploy a text model or a data parsing agent without requiring explicit tool selection.
11 more episodes mention GPT-5.4.
Create an account to see the whole feed, search across every transcript, and follow the entities you care about.