Jul 31, 2026 · 36 min · 12 segments
China's latest AI breakthrough is putting new pressure on OpenAI, Google, and the entire AI industry. This week on 4 Guys Talking About AI, Greg and John examine how low-cost open-weight AI models…
John GerardHostGreg SterlingHostOkay, so John, um, maybe you can take, take us through, uh, what happened with the, uh, Kimi K2, Kimi K3 Chinese open, open weight model that, that landed this week.

Uh, boy, and listening to that list of things you were covering today, we've, we've got some breadth here.

So, um, so, uh, he- here's the, the, the summary on the open weight models and the Chinese, uh, models that have come out.

I think somewhat unsurprisingly, the, uh, the models that are, that are coming out a- at each iteration or generation are surprising everybody in terms of how good they are and how much better they, uh, they, they are than expected.

We've seen this just time and time again, where every time a new model comes out, uh, either from the US company or the Chinese side, there's all of this kind of mind-blown, uh, responsiveness.

I think at some point we have to just, uh, uh, realize that this is, this is, uh, incredible and, and, and, uh, going to continue to be so.

The implications are, are, are pretty interesting, and what may be going on behind the scenes, also pretty interesting.


Um, do these models offer an alternative at, you know, 10%, 5%, or even less the cost of using, uh, Claude's APIs? And this is a, a, a, a, a general, um, uh, you know, a sort of a concern for the US AI companies who have had to have these pretty aggressive API prices in order to start getting at least some revenue.

So, you know, what's going on here? Well, it could be, uh, potentially very disruptive.

Here's a nuance that I think gets lost that will have to be solved, but when you're talking about models that are 5% as ex- as expensive, there's a lot of incentive to solve them.


And in fact, there's a lot more than just the model that is under the hood that makes these things go.

On the coding side, it's called a harness, and a harness is really the set of instructions that tells the model how to approach coding projects.


I, I, you know, anecdotally happen to think that, you know, more than half of the capabilities that something like Claude Code has have to do with the harness as opposed to the underlying model, and that may become even more important as these models achieve parity.
Okay, so John, um, maybe you can take, take us through, uh, what happened with the, uh, Kimi K2, Kimi K3 Chinese open, open weight model that, that landed this week.

Uh, boy, and listening to that list of things you were covering today, we've, we've got some breadth here.

So, um, so, uh, he- here's the, the, the summary on the open weight models and the Chinese, uh, models that have come out.

I think somewhat unsurprisingly, the, uh, the models that are, that are coming out a- at each iteration or generation are surprising everybody in terms of how good they are and how much better they, uh, they, they are than expected.

We've seen this just time and time again, where every time a new model comes out, uh, either from the US company or the Chinese side, there's all of this kind of mind-blown, uh, responsiveness.

I think at some point we have to just, uh, uh, realize that this is, this is, uh, incredible and, and, and, uh, going to continue to be so.

The implications are, are, are pretty interesting, and what may be going on behind the scenes, also pretty interesting.


Um, do these models offer an alternative at, you know, 10%, 5%, or even less the cost of using, uh, Claude's APIs? And this is a, a, a, a, a general, um, uh, you know, a sort of a concern for the US AI companies who have had to have these pretty aggressive API prices in order to start getting at least some revenue.

So, you know, what's going on here? Well, it could be, uh, potentially very disruptive.

Here's a nuance that I think gets lost that will have to be solved, but when you're talking about models that are 5% as ex- as expensive, there's a lot of incentive to solve them.


And in fact, there's a lot more than just the model that is under the hood that makes these things go.

On the coding side, it's called a harness, and a harness is really the set of instructions that tells the model how to approach coding projects.


I, I, you know, anecdotally happen to think that, you know, more than half of the capabilities that something like Claude Code has have to do with the harness as opposed to the underlying model, and that may become even more important as these models achieve parity.
The rest of this transcript — segmented and speaker-labeled, so you land on the exact moment something was said
Search every transcript — by keyword, by phrase, or by meaning, across every show Radar indexes
Trends — what is surging across podcasts, measured against its own baseline
Alerts — when a name you follow appears in a newly indexed episode
No account is needed to search Radar.