
Gavin Uberti
5
APPEARANCES
5
PODCASTS
024
DEC 30
JAN 6
JAN 13
JAN 20
JAN 27
FEB 3
FEB 10
FEB 17
FEB 24
MAR 3
MAR 10
MAR 17
MAR 24
MAR 31
APR 7
APR 14
APR 21
APR 28
MAY 5
MAY 12
MAY 19
MAY 26
JUN 2
JUN 9
JUN 16
JUN 23
JUN 30
JUL 7
JUL 14
JUL 21
JUL 28
AUG 4
AUG 11
AUG 18
AUG 25
SEP 1
SEP 8
SEP 15
SEP 22
SEP 29
OCT 6
OCT 13
OCT 20
OCT 27
NOV 3
NOV 10
NOV 17
NOV 24
DEC 1
DEC 8
DEC 15
DEC 22
DEC 29
JAN 5
JAN 12
JAN 19
JAN 26
FEB 2
FEB 9
FEB 16
FEB 23
MAR 2
MAR 9
MAR 16
MAR 23
MAR 30
APR 6
APR 13
APR 20
APR 27
MAY 4
MAY 11
MAY 18
MAY 25
JUN 1
JUN 8
JUN 15
JUN 22
JUN 29
JUL 6
JUL 13
JUL 20
JUL 27
AUG 3
AUG 10
AUG 17
AUG 24
AUG 31
SEP 7
SEP 14
SEP 21
Aug 19, 2026
July Fed Minutes, Etched CEO Interview, Latest Ultra-Wealthy Investments 8/19/26
35:26
35:32
35:40
35:48
35:54
36:07

Gavin UbertiGUEST
Well, the thing about GPT-3 is that it was much smarter than its predecessors, largely by virtue of being way bigger.

Gavin UbertiGUEST
And that made me very confident that models would keep getting smarter as time went on.

Gavin UbertiGUEST
Now, when people are talking about training, training, training, people didn't realize the cost of training is kind of a fixed price.

Gavin UbertiGUEST
But if you wanna go ahead and serve to many, many billions of people, the inference cost is what scales.

Gavin UbertiGUEST
We said, if we're gonna go do one thing really well, we're gonna build the world's best inference solution.
Kelly EvansHOST
So what is the new idea that you're working on? And what does it do to make clients like Jane Street or anybody else better?
America's Tech Wishlist, Comcast Splits in Two, Tristan Thompson in the Ultradome | Zach Laberge, Gavin Uberti, Eren Bali, Larsen Jensen, Ricky Rosa, Andrew Rea, Michael Anderson, Tristan Thompson, Grant Gregory & Ian Rountree
60:05
60:11
60:17
69:18
69:32

Gavin UbertiGUEST
You can't really run a GPU more than around 50% of what it could theoretically do, or it'll melt.

Gavin UbertiGUEST
So we're introducing a new technology today called low voltage inference to try to solve this problem.

Gavin UbertiGUEST
And what that is, is we bring the voltage of the chip down dramatically, which allows us to have way, way better efficiency in terms of how much power is drawn per unit of math, and thus fit way, way more flops onto the chip.
9 MINS LATER

Gavin UbertiGUEST
When you think about economies of scale, it all just comes down to how much is there to go serve? If the market's relatively small, you justify a small factory, but not some gigantic mega cluster.

Gavin UbertiGUEST
And with what we're seeing right now with these many, many trillion parameter models, with these, uh, quadrillion token, uh, demands that we're seeing that are increasing every month, there has never been a better time to go ahead and invest in those economies of scale.
Chip Stocks on Track for Best Quarter Ever
S
35:15speaker_17ADVERTISER
Complete disclosures available at public.com/disclosures.
Etched - Building AI Hardware to Make Inference Faster and Cheaper - [Invest Like the Best, EP.480]
6:00
6:11
6:19
6:44

Rob LockettGUEST
That's, like, a very simple one, but there's many more that you get, twenty percent here, fifty percent there, two X here, and these compound to a system that can be radically better for inference.

Gavin UbertiGUEST
There are some folks who went purely on heuristics of, "Hey, young founders, they claim they can go beat the biggest company in the world on performance.

Gavin UbertiGUEST
And there is no thing you could go say to me that would make me change my mind." But there's also people out there who are, of course, skeptical, but are willing to go ahead and say, "I'll spend the time, I'll do the work, and is this actually possible?" Like, for example, one of our earliest, earliest supporters was Mark Ross, and Mark was a very prestigious semiconductor expert.

Gavin UbertiGUEST
And when we met him, we were just a couple of guys in a dorm room, and we came to him and say, "Hey, we want to go build hardware for inference.
43 MINS LATER
49:53
When will that just be something that AI does en-entirely as well? Are humans still the best kernel engineers? Are they doing it with the assistance of AI systems? Like, how far down will humans still be in the loop of designing these things? Like, when will that go away?
23andMe BANKRUPT, Saylor Still Buying Bitcoin, + Founder Interviews | E2102
23:18

Alex WilhelmHOST
So I want you to tell me why you picked that thesis and also how the progress has been to this date.
G
23:24Gavin UbertiGUEST
Right now, it is clear that AI inference is going to be a massive, massive market and there are already a number of chips on the market like NVIDIA's GPUs and Google's TPUs that do a pretty good job of this.
G
23:37Gavin UbertiGUEST
You can go run code on an NVIDIA GPU or a Google TPU to run many different kinds of models.
G
23:43Gavin UbertiGUEST
Convolutional networks like ResNets, Transformers of course, LSTMs, RNNs, whatever wacky thing comes out of, you know, neural network factories.
G
23:53Gavin UbertiGUEST
Because they're so flexible, because so much space on these chips is spent on caches and control for the logic, only a very small fraction of the die is actually spent on the math blocks, the MatMuls that do the work to run the AI model.
9 MINS LATER

