Skip to main content

Lily Chen

Integrative veterinarian ("The Unicorn Vet"), founder of Integrative Pet Wellness Center in Rolling Hills Estates, CA, and host of the My Dog Is Better Than Your Dog podcast.

Jun 5, 2026

3:01
Like, I mean, I guess on the human side, social engineering, "Send me, send me through this, this gift card, and you could win millions." And I guess the LLM equivalent of that is some kind of amazing sparkly piece of data that it really needs.
3:15
Yeah.
3:16
And so three years ago, when LLMs just started to become popular, I was, like, looking to this topic.
3:23
And at that time, tricking the AI to, tricking the AI into believing something is quite easy.
3:29
You basically give them an emotional blackmail saying that, "Oh, you must believe in this," or maybe, "My grandmother will be very sad." Something like this is enough to trick the AI.
3:39
But we have gone through three years of AI safety and security, so in modern LLMs, they are not so vulnerable.
3:47
So what we do is, in modern language models, there's this thing called attention, and, like, when they're trying to generate the next token, they consider the input, but not equally.

5 MINS LATER

9:08
I guess not protect against it because it's a cat-and-mouse game, but, you know, how, how can companies who might be concerned against this start to look at how to harden their workflows?

We value your privacy

We use cookies to understand how you use our platform and to improve your experience. Click “Accept All” to consent, or “Decline non-essential” to opt out of non-essential cookies. Read our Privacy Policy.