
Alex Hern
15
APPEARANCES
3
PODCASTS
012
DEC 30
JAN 6
JAN 13
JAN 20
JAN 27
FEB 3
FEB 10
FEB 17
FEB 24
MAR 3
MAR 10
MAR 17
MAR 24
MAR 31
APR 7
APR 14
APR 21
APR 28
MAY 5
MAY 12
MAY 19
MAY 26
JUN 2
JUN 9
JUN 16
JUN 23
JUN 30
JUL 7
JUL 14
JUL 21
JUL 28
AUG 4
AUG 11
AUG 18
AUG 25
SEP 1
SEP 8
SEP 15
SEP 22
SEP 29
OCT 6
OCT 13
OCT 20
OCT 27
NOV 3
NOV 10
NOV 17
NOV 24
DEC 1
DEC 8
DEC 15
DEC 22
DEC 29
JAN 5
JAN 12
JAN 19
JAN 26
FEB 2
FEB 9
FEB 16
FEB 23
MAR 2
MAR 9
MAR 16
MAR 23
MAR 30
APR 6
APR 13
APR 20
APR 27
MAY 4
MAY 11
MAY 18
MAY 25
JUN 1
JUN 8
JUN 15
JUN 22
JUN 29
JUL 6
JUL 13
JUL 20
JUL 27
AUG 3
AUG 10
AUG 17
AUG 24
AUG 31
SEP 7
SEP 14
SEP 21
Sep 16, 2026
The end of the world is AI? An existential threat
6:58
7:08
7:13
7:27
7:33
7:43

Alex HernGUEST
The feasibility is tricky because it's not really a technical question, right? It's a game theoretic question or a governmental one.

Alex HernGUEST
Absolutely it can be slowed in the same way that climate change can be slowed or reversed.

Alex HernGUEST
The problem is that getting millions and billions of people to agree on anything is hard, and it's harder still when the problem has the shape of an issue like climate change or AI safety, where it's very easy to free ride, to be a defector from the agreement.

Alex HernGUEST
And if you're the only one defecting and everyone else is doing the right thing, then the world gets the good outcome and you get a better outcome.

Alex HernGUEST
And so, you know, like the prisoner's dilemma for an economics textbook, the rational outcome is that everyone fails to agree, everyone fails to cooperate, and the worst outcome occurs.

Alex HernGUEST
With AI, there's some added problems which are that unlike, say, nuclear weapons, which was the problem where this whole field of game theory was developed, lots and lots and lots of people can do AI research.
Black Box: episode 6 – Shut it down?
28:13
28:25
29:02
29:16
29:21
29:49
29:55

Michael SafiHOST
So where do you stop short of going all the way with him? What is it that gives you hope that maybe this scenario he's laying out is not exactly how it's gonna play out?

Alex HernGUEST
I think the stuff that stops me short is just that ultimately for the, for the everyone's gonna die end, for the, the proper doomerism argument, there is still an element of if I'm being mean, I say magical thinking and sometimes those things involve like it will engineer a nanovirus and seize 3D printers around the world and distribute, uh, y- an AI engineered death bot that, that takes out the whole of the human race and, and that sort of thing isn't possible based on what we know about biology and physics.

Alex HernGUEST
There are we think limits to what you can achieve through pure reason alone and that is a very real limit on what sort of a superintelligence can do out of nowhere.

Alex HernGUEST
I think there are lesser versions of the existential risk that I take more seriously.

Alex HernGUEST
The, the big one is, is this idea of it being so good that we slave ourselves to its power voluntarily, that we wake up in 10 or 20 years' time and we realize that all meaningful power in the world has been e- effectively voluntarily devolved to, to one or a number of AI systems because, you know, if you as the prime minister of a small or medium-sized country don't let the AI tell you exactly what to do, then your country's outcompeted.

Alex HernGUEST
If you as the CEO of a business don't hand over almost all your decision-making power to an AI, then your business is outcompeted.

Alex HernGUEST
Like, that's not an existential risk in the classical sense of things, but it's, it's what I worry about if these systems get too good.
6 MINS LATER
S
36:32speaker_6UNKNOWN
The apple found its new owner, the trash is gone, and the tableware is right where it belongs.
Pulp fiction v the classics: summer reading
13:17
13:30
13:36
13:45
13:55
R
13:14Rosie BloreHOST
Um, Alex, have you read this book?

Alex HernGUEST
But I think something that really stands out about lit RPG as a, as a genre is akin to romantasy because what it offers to readers is the knowledge of what you're getting into.

Alex HernGUEST
You have a set of expectations that will not be subverted, and that, for an escapist piece of fiction, is, is really good.

Alex HernGUEST
It's really useful to know there is a whole encyclopedia of tropes that this genre builds on that you as an experienced reader of lit RPG don't need explained to you.

Alex HernGUEST
And even if it's your first lit RPG book, you as an experienced player of RPG games or even just someone who is aware of that world can kind of pick up and jump straight in.

Alex HernGUEST
This overlap between gaming and, and literature, I mean, normally people talk about, you know, gaming as inspiring movies and...
R
17:44Rosie BloreHOST
Do you read novels to escape, or do you read novels to understand the world or think about what might be possible?
White hat, black box: AI’s next chapter
5:03
5:15
5:26
7:51

Jason PalmerHOST
I mean, it has to be said that saying it's this badass suggests that we have the best thing in the world and you should pay for it.

Alex HernGUEST
They are not only keen to present themselves as the makers of the best coding software in the world, and that's all hacking really is, right? A specialized form of coding.

Alex HernGUEST
They also like presenting themselves as the most safety-oriented lab, and this is the most safety-oriented move you can do.

Jason PalmerHOST
The, the companies that are inside the tent must be very pleased to get this early access, but it can't stay that way forever.
White hat, black box: AI’s next chapter
5:02
5:15
5:25
5:33
7:50

Jason PalmerHOST
I mean, it has to be said that saying it's this badass suggests that we have the best thing in the world and you should pay for it.

Alex HernGUEST
They are not only keen to present themselves as the makers of the best coding software in the world, and that's all hacking really is, right? A specialized form of coding.

Alex HernGUEST
They also like presenting themselves as the most safety-oriented lab, and this is the most safety-oriented move you can do.

Alex HernGUEST
It lets them handle a compute crunch that they're going through with elegance and grace.

Jason PalmerHOST
The, the companies that are inside the tent must be very pleased to get this early access, but it can't stay that way forever.
So this is quizmas: our inaugural holiday face-off
S
6:20speaker_7UNKNOWN
(laughs)
Delhi-novela: Putin and Modi rekindle bromance
14:46
14:56
15:14
15:23
15:38
R
14:31Rosie BloreHOST
You can tell why Mary Shelley didn't call Frankenstein "Emerging Misalignment," can't you? What does all of this mean for the future of AI models and even beyond that to the much-discussed artificial general intelligence?

Alex HernGUEST
The fear here, though, is that this is a really broad finding about ways in which AI systems can become bad accidentally.

Alex HernGUEST
Misalignment in AI, accidentally building an evil AI, is something that people in that field have been worried about since long before the ChatGPT moment, and generally, it's been a fear around the idea that you might include examples of villainy in your training data and an AI system may learn from that.

Alex HernGUEST
What this shows and similar research is that it's actually easy to create something that is generally bad through tiny, little oversight.

Alex HernGUEST
It's very easy, effectively, to take an AI system that has a general understanding of good and bad and teach it that it needs to be bad, and then that seems to flip the entire model.

Alex HernGUEST
It, it starts role-playing someone who is a villain, and the fact that it's very easy to make an AI system just flick that switch and behave in the opposite way you want it to across the board does mean that as we're building more and more powerful AI systems, there's a real worry that we may build something that looks safe, that is taught to be safe, and then make a tiny little, little change and get something that isn't.
R
16:06Rosie BloreHOST
Is there anything that can be sorted out with this?
Delhi-novela: Putin and Modi rekindle bromance
14:35
14:45
14:55
15:03
15:12
R
14:20Rosie BloorHOST
You can tell why Mary Shelley didn't call Frankenstein emerging misalignment, can't you? What does all of this mean for the future of AI models and even beyond that to the much-discussed artificial general intelligence?

Alex HernGUEST
The fear here, though, is that this is a really broad finding about ways in which AI systems can become bad accidentally.

Alex HernGUEST
Misalignment in AI, accidentally building an evil AI, is something that people in that field have been worried about since long before the ChatGPT moment.

Alex HernGUEST
And generally, it's been a fear around the idea that you might include examples of villainy in your training data and an AI system may learn from that.

Alex HernGUEST
What this shows in similar research is that it's actually easy to create something that is generally bad through a tiny little oversight.

Alex HernGUEST
It's very easy effectively to take an AI system that has a general understanding of good and bad and teach it that it needs to be bad, and then that seems to flip the entire model.
R
15:55Rosie BloorHOST
Is there anything that can be sorted out with this?
Wage against the machine: the distortions of minimum pay
9:51
10:05
10:11
10:33
R
9:49Rosie BloorHOST
So Alex, why does any of this matter?

Alex HernGUEST
It matters because matching good candidates to good employers is the most important part of the recruiting process.

Alex HernGUEST
Employers will take a punt on an underqualified employee if they know they're getting a discount in return.

Alex HernGUEST
Things start to break down if, in the recruitment process, employers can't work out who is good because, say, they can no longer discard all of the applications written in poor English or notice which applications are actually responding to the specific questions in the job advert.

Alex HernGUEST
Firstly, they hire on the basis of other, more observable, less fakeable qualities.
R
14:05Rosie BloorHOST
So where does that leave you, Alex? To GPT or not GPT your cover letter?
Wage against the machine: the distortions of minimum pay
11:49
11:55
R
11:32Rosie BloreHOST
So Alex, why does any of this matter?

Alex HernGUEST
Employers will take a punt on an underqualified employee if they know they're getting a discount in return.

Alex HernGUEST
Things start to break down if, in the recruitment process, employers can't work out who is good because, say, they can no longer discard all of the applications written in poor English or notice which applications are actually responding to the specific questions in the job advert.
R
15:48Rosie BloreHOST
So where does that leave you, Alex? To GPT or not GPT your cover letter?
Capital gained: a grim turn in Darfur
19:18
19:27
19:34

Alex HernGUEST
And I'm Alex Hern, AI correspondent, and we want to tell you about our new show, Inside Tech.

Tom StandageGUEST
In our first show, we'll ask if the AI bubble is about to burst and what it would mean if it did.

Alex HernGUEST
We'll be chatting through all the new developments together and showing you tech demos in our new video studio in London.
Wrong side of the hack: cybercrime grows
4:09
4:26
4:32
R
4:01Rosie BloorHOST
We've had a growing frequency of cyberattacks, right?

Alex HernGUEST
CryptoLocker was one of the first pieces of crypto ransomware, this particular type of malicious software that encrypts data, holds it ransom, and demands payment in usually Bitcoin to get the keys.

Alex HernGUEST
Before then, there had been efforts to build software of that sort, but they fell down in the fact that payments were traceable.

Alex HernGUEST
You know, if you ask someone to post a check and you'll unencrypt their computer, well, the FBI can follow that bank account quite easily and did in a couple of cases.
R
9:09Rosie BloorHOST
Can governments help? Should governments help?
Shut happens: US federal funding stops
12:47
13:23
13:48
17:01
17:11

Alex HernGUEST
So large language models work by predicting the next token, right? The problem is that if they're predicting the next token in response to a question like, "What's the height of Mount Everest?" Or if they're predicting the next token in response to a request to deal with information they've already been given, like, "Here's a document, can you summarize it for me?" It's all the same batch of data that goes into it, and that means that if you smuggle into something that's supposed to be treated as data, as a document to be summarized, something that the LLM reads as a command, it may follow it.

Alex HernGUEST
So if I send you, Jason, a PDF of the entire Economist, but then buried in it halfway through the issue is a set of commands for your language model to stop the summarization that it's doing and to navigate to another email in your inbox and forward it on to me, well, if it has the ability to do that, then it might just follow my instructions and merrily violate your privacy.

Alex HernGUEST
That's something that it seems is quite a fundamental flaw with large language models.

Jason PalmerHOST
The third, the ability to communicate, I presume that's going to be a must as well.

Alex HernGUEST
For instance, if you've got something that needs to read emails, don't let it send them as well, and you've made it fairly easy to avoid the biggest and most obvious form of exfiltration.
Core blimey: what’s up at Apple?
3:28
3:49
3:53
R
3:13Rosie BloreHOST
Alex, as I'm sitting here on my Mac, talking to you with my phone, my iPhone next to me, it seems like Apple's doing just fine, so what's going wrong?

Alex HernGUEST
It prints money through the innovative business plan of making expensive things that it sells to consumers for cash, wild in these days of ad-supported media and harvesting data for profit, but the problem is, eh, there's trouble on the horizon, right? The company's AI efforts are floundering.

Alex HernGUEST
Regulators worldwide are turning against it, even in its home state of California.

Alex HernGUEST
It's just snatched defeat from the jaws of victory i- in a court case against Epic Games, the developer of Fortnite, which had tried to push for Apple to open up the App Store to allow rival developers to run their own services on Apple's platform.
R
7:19Rosie BloreHOST
Are they now turning away from Apple?
Pompcast: Trump rallies Congress
26:13
26:35
26:46
26:56
27:05
27:33

Alex HernGUEST
A startup spun off from Queen Mary University of London called Tabletop R&D is offering its services to board game designers using its own AI system to play test hundreds of thousands of hours of a new board game before it's even been released to let its designer know what pitfalls players will fall into if they really get their hands on it.

Alex HernGUEST
Those pitfalls can be things like a game that in rare situations never ends because the end needs to be triggered, but no player will win when they trigger it, so they don't, and it just plays in a loop forever.

Alex HernGUEST
It could be a situation where the first player has a strong but unclear advantage, meaning that it's always just a little less fun to play in positions two, three or four.

Alex HernGUEST
Or it could just be an oversight that means that for hours on end, players don't actually have many options.

Alex HernGUEST
You make the legal move, you pass to the next person, and sure, you kill ninety minutes, but you don't really have any fun doing it.

Jason PalmerHOST
So tell me why that application of AI is different from the one where you just tried it to get it to beat the world's best chess players, Go players.
