K
Kaitlyn Zhou
1
APPEARANCES
1
PODCASTS
012
DEC 30
JAN 6
JAN 13
JAN 20
JAN 27
FEB 3
FEB 10
FEB 17
FEB 24
MAR 3
MAR 10
MAR 17
MAR 24
MAR 31
APR 7
APR 14
APR 21
APR 28
MAY 5
MAY 12
MAY 19
MAY 26
JUN 2
JUN 9
JUN 16
JUN 23
JUN 30
JUL 7
JUL 14
JUL 21
JUL 28
AUG 4
AUG 11
AUG 18
AUG 25
SEP 1
SEP 8
SEP 15
SEP 22
SEP 29
OCT 6
OCT 13
OCT 20
OCT 27
NOV 3
NOV 10
NOV 17
NOV 24
DEC 1
DEC 8
DEC 15
DEC 22
DEC 29
JAN 5
JAN 12
JAN 19
JAN 26
FEB 2
FEB 9
FEB 16
FEB 23
MAR 2
MAR 9
MAR 16
MAR 23
MAR 30
APR 6
APR 13
APR 20
APR 27
MAY 4
MAY 11
MAY 18
MAY 25
JUN 1
JUN 8
JUN 15
JUN 22
JUN 29
JUL 6
JUL 13
JUL 20
JUL 27
AUG 3
AUG 10
AUG 17
AUG 24
AUG 31
SEP 7
SEP 14
SEP 21
Aug 10, 2026
A Scientific Deep Dive into Overconfident LLMs: Interview with Kaitlyn Zhou (Cornell)
7:24
7:36
8:04
8:16

Kaitlyn ZhouGUEST
I think when we started the project, we wanted to know how language models were going to generate these epistemic markers and how they were going to express certainty and uncertainty.

Kaitlyn ZhouGUEST
But before we could even answer that question, we first had to deal with the idea of do language models even understand what these markers even mean? Do they even understand what epistemic markers are? And so we did a much easier experiment to start off, which was that first paper in the series of three papers we wrote, which was if I prompt a language model with saying, What is the capital of France? I'm certain the answer is blank.

Kaitlyn ZhouGUEST
We wanted to know, could it recognize that if it says, I'm certain the answer is blank versus I'm not sure the answer is blank, that when it says, I'm certain it's more likely to produce the correct answer.

Kaitlyn ZhouGUEST
And when we lead it with, I think the answer is, it's more likely to produce the incorrect answer.
21 MINS LATER
