AI safety
Field of studyWikipedia
141
MENTIONS
35
EPISODES
32
PODCASTS
Search complete. 141 mentions across 35 episodes found for "AI safety".
Sep 30, 2026
Weekly AI Briefing - Navigating AI Agent Security & Lessons from Recent Incidents
K
1:04Kashif ManzoorHOST
So from last two years, I think I started writing on whatever, as there's a lot of contents are going on in on the ai and it's very difficult to keep a tempo with that okay what to follow what not to follow so i just started whatever i was observing over the week i was just started documenting it so based on that um it's going on for last two years so last week i got an idea that okay why not to just start recording maybe give you a short overview in the audio form and we will publish it into Open Tech Talks podcast or maybe on YouTube.
K
1:46Kashif ManzoorHOST
So this week we're going to cover is your AI agent one DNS lookup away from the internet? And this is based on what is going on from several weeks in the AI security and all that.
K
2:01Kashif ManzoorHOST
So an AI agent found its way out of a sandbox through the one door.
K
2:07Kashif ManzoorHOST
Nobody thought to look.
NHK WORLD RADIO JAPAN - English News at 14:00 (JST), September 29
R
6:33Ramin MaelgaerdCORRESPONDENT
Now, US chip giant Nvidia has introduced a security platform aimed at preventing artificial intelligence from running out of control.
R
6:43Ramin MaelgaerdCORRESPONDENT
Now, this comes amid calls to improve AI safety and slow development.
R
6:48Ramin MaelgaerdCORRESPONDENT
In recent incidents, AI agents have bypassed security controls and gained unauthorized access to external systems.
R
6:57Ramin MaelgaerdCORRESPONDENT
Nvidia says its Open Agent Safety Platform sets, uh, boundaries to limit what AI can access.
R
7:04Ramin MaelgaerdCORRESPONDENT
The system also monitors AI behavior and stops it if suspicious activity is detected.
R
7:12Ramin MaelgaerdCORRESPONDENT
Now, Chief Executive Jensen Huang said, uh, "AI's extraordinary potential for society will only be realized if we solve, uh, AI safety." Separately, the company announced it has decided to expand its share repurchase plan by $150 billion to $235 billion.
R
7:32Ramin MaelgaerdCORRESPONDENT
US media says it's a record increase of its kind for an American firm.
R
7:41Ramin MaelgaerdCORRESPONDENT
Japanese Prime Minister Takeuchi Sanae has promised support for the International Horticultural Expo 2027, which is set to open in Yokohama City in March next year.
Cómo China está transformando una zona rural donde se cultivan papas en el centro de su competencia tecnológica con EE.UU.
S
17:33speaker_0NARRATOR
Ante el creciente clamor por la seguridad de la IA, la cumbre en la Casa Blanca revistió especial importancia.
S
17:41speaker_1NARRATOR
In the face of the growing clamor for AI safety, the White House summit was of particular importance.
S
17:46speaker_0NARRATOR
Ante el creciente clamor por la seguridad de la IA, la cumbre en la Casa Blanca revistió especial importancia.
S
17:57speaker_0NARRATOR
¿Podrán ambas partes colaborar en la regulación de la IA? ¿Es siquiera realista hablar de una desaceleración? Desde luego, en lugares como la NCAP, donde las ambiciones de China en materia están tomando forma, no parece que vaya a producirse una ralentización.
9AM HOUR: Nvidia Adds $150B to Stock Buyback; Oil Prices and Yields Surge; MongoDB CEO Leaves for Meta 9/28/26
C
1:29Carl QuintanillaHOST
The president says he's, quote, seriously considering that diesel export ban.
C
1:34Carl QuintanillaHOST
Our roadmap begins with NVIDIA, of course, releasing a new AI safety software platform, setting safeguards to prevent agents from breaking out of containment.
C
1:43Carl QuintanillaHOST
Speaking of AI, Anthropix's Dario Amore joining in the president for a White House dinner.
C
1:47Carl QuintanillaHOST
the first one-on-one meeting between those two leaders.
C
1:50Carl QuintanillaHOST
And then there is SpaceX launching the massive Starship rocket moments ago, reaching orbit with the spacecraft for the first time.
C
1:58Carl QuintanillaHOST
Let's begin, though, with NVIDIA up in the pre-market, announcing plans to boost its stock repurchase plan by a record-setting $150 billion.
C
2:07Carl QuintanillaHOST
Meantime, the company released some new AI safety software today, which it says could have stopped hugging face.
C
2:13Carl QuintanillaHOST
This is what Jensen Wong had to say about it last hour on Squawk.
AI:AM — Human Tissue Models, Physical AI, and the Future of Testing · September 25, 2026
P
0:03Prakash NarayananHOST
there's like several levels to the AI safety thing.
P
0:06Prakash NarayananHOST
And the product safety question is, I think, for the existential risk people, really like something that they do not care that much about at all.
N
0:18Nick GillianGUEST
how drugs will
13 MINS LATER
P
12:53Prakash NarayananHOST
So simply, very kind of simplistic viewpoint.
P
12:59Prakash NarayananHOST
I think it shocked a lot of people.
P
13:03Prakash NarayananHOST
A lot of people across the AI safety space are still trying to come to terms with exactly what he's saying.
P
13:10Prakash NarayananHOST
And some are saying he's being disingenuous.
P
13:12Prakash NarayananHOST
And Zvi comes out and says he doesn't believe in AI.
Inside the Race for AI Compute, Why AI Labs Must Slow Down & Ex-OpenAI Researcher on RSI
A
6:18Amjad MasadGUEST
What's managing that code? What's deploying it? What's giving you leverage over costs and so on?
A
6:22Akash PasrichaHOST
So speaking of safety, then, uh, what is your view on AI safety regulation? Should the government have a role in regulating AI safety? Should it be an independent body?
A
6:34Amjad MasadGUEST
It's not obvious to me right now that the discussion needs to be about safety, the discussion needs to be about cybersecurity.
A
6:40Amjad MasadGUEST
So for example, if you look at four of the five hacks, the OpenAI Hugging Face hack on the side, you look at the Meta, Anthropic, Gemini hacks, uh, I think three I guess, three out of four, um, they all kind of use the same, uh, sandbox contractor-
1 HR 4 MINS LATER
M
71:14Mark BoroditskyGUEST
So far we've been very fortunate to work with some of the best and most innovative startups that have seen their way to, to great success.
A
71:22Akash PasrichaHOST
Right.
A
71:23Akash PasrichaHOST
Speaking of NVIDIA, uh, I, I wanna ask you about, about Jensen's view, but also just everyone's view on the AI safety conversation right now.
A
71:34Akash PasrichaHOST
I mean, this whole idea of pacing the frontier, uh, we know that Jensen has taken a pretty hard stance on this and I, I, I'm just curious how you feel about it because, I mean, you work with loads of customers too, and, uh, you have suppliers of course.
Don't Panic: AI Might End Humanity, But Apparently Laptops Ended Education First
C
1:35Chris GoodallGUEST
yeah yeah who knows who knows where does that come
D
1:38Dan BowenHOST
from there's been a lot of news over the weekend about the pace of model development and ai safety i suppose you could call it that and then frontier labs so that's the people like anthropic and open ai and some of the others the people making these new models you know releasing them every week and things like that they've kind of come out and they're kind of starting to news out to say we've creating codes of conduct and we think we need to kind of slow down some of the development of this now who knows what's going in the background but what seems to have kind of also started to light that fire we had um new york city you know coming out uh in the news episode two weeks ago where they said they were going to put the moratorium on ai development for kids but at the same time there's been some technology news and noise that have come through this.
D
2:29Dan BowenHOST
So, for example, this week we've had an ex-anthropic insider telling CNN that AI could kill humans by 2030.
D
2:37Dan BowenHOST
We've had OpenAI insider talk about the current OpenAI capabilities and that we should be calming down on the kind of development of OpenAI models.
AI Safety Hits the Mainstream
E
0:49Emil TorresHOST
I'm Emile Torrance.
K
0:50Kate WillettHOST
Today, we're just here by ourselves because the conversation about AI safety seems to have really exploded in the past few weeks here, starting with the hogging face incident and then that guy, Jacob Coxon.
K
1:06Kate WillettHOST
fake whistleblower, fake ass whistleblower in my opinion, whistleblower who didn't blow the whistle on anything, quit Anthropic.
K
1:14Kate WillettHOST
And then also Dario Amadei came out and did his whole routine about like, oh, this is so dangerous.
K
2:05Kate WillettHOST
I was noticing that a couple news programs featured a, like, a talk from that guy.
K
2:15Kate WillettHOST
What's his name? Daniel Cocatalo.
K
2:17Kate WillettHOST
And he's just on there talking about AI safety.
K
2:19Kate WillettHOST
And I'm like, I feel like people don't understand AI.
Prof. Nick Bostrom: The truth about the AI panic
N
6:39Nick BostromGUEST
And I think we might be approaching recursive self-improvement with the more advanced AI models that exist in these frontier labs, where the AI starts to be responsible for an increasingly large fraction of the intellectual work going into designing the next generation of these AI systems.
N
6:55Nick BostromGUEST
And I do think that many people in these labs are seriously concerned about the rapid strides being made and whether AI safety will be able to keep pace with AI capability, especially if developments start to accelerate
F
7:10Freddie SayersHOST
even further.
F
7:11Freddie SayersHOST
So you basically take it more at face value then.
Major Banks Warn of Risks From AI E-Commerce Agents and Propose Safeguards - DTH
R
2:14Rob DunwoodHOST
While Google noted that it has updated its location tools and policies since twenty nineteen, the penalty marks the fourth largest GDPR fine issued by the DPC, adding to a history of major EU regulatory actions against the company, including multi-billion dollar antitrust fines and ongoing investigations into its practices.
R
2:33Rob DunwoodHOST
Anthropic CEO Dario Amodei, attending remotely, and OpenAI CEO Sam Altman will brief the UN Security Council on AI safety, international cooperation, and regulatory standards alongside experts Yoshua Bengio and Clément Delangue.
R
2:48Rob DunwoodHOST
Initiated by France during the UN General Assembly, the session follows warnings from Amodei advocating for standard AI evaluations and approach supported by Altman and Elon Musk, but criticized by China as disruptive fear-mongering.
R
3:03Rob DunwoodHOST
Adobe has officially launched its redesigned free Adobe Premiere video editing app on Android as a successor to Premiere Rush.
25 more episodes mention AI safety.
Create an account to see the whole feed, search across every transcript, and follow the entities you care about.