
AlphaGo Zero
SoftwareWikipedia
3
MENTIONS
2
EPISODES
2
PODCASTS
Search complete. 3 mentions across 2 episodes found for "AlphaGo Zero".
Sep 22, 2026
#241: Pacing the Frontier Gets Political, Why AI Labs Could Keep the Best Models for Themselves & Introducing the AI Transformation Blueprint
P
31:49Paul RoetzerHOST
In 2016, he said, "AlphaGo beat Lee Sedol in a milestone for AI, but key to that was AI's ability to, quote, 'ponder' for one minute in, uh, before each move.
P
32:01Paul RoetzerHOST
How much did that improve it? For AlphaGo Zero, it's the equivalent of a scaling pre-training by 100,000x." Then he said, "All those prior methods are specific to that game.
P
32:13Paul RoetzerHOST
But if we can discover a general version, the benefits could be huge.
P
32:17Paul RoetzerHOST
Yes, inference may be 1,000 times slower and more costly," meaning let- letting the model think before it responds.
Inside the World of Reinforcement Learning
S
18:55speaker_5HOST
Mm-hmm.
S
18:56speaker_4HOST
But the later versions, like AlphaGo Zero, they didn't learn from humans at all.
S
19:01speaker_5HOST
And that represents the ultimate paradigm shift in the field.
S
19:04speaker_5HOST
AlphaGo Zero started with completely blank slate.
S
19:06speaker_5HOST
It only knew the rules of how a stone could be placed.
S
19:09speaker_5HOST
It generated its own training data by playing millions of games against itself.