Skip to main content
Diffusion model

Diffusion model

Search complete. 14 mentions across 3 episodes found for "Diffusion model".

Sep 8, 2026

speaker_1NARRATOR
1:48
Speaking of discovery, let's talk about the technical side.
speaker_1NARRATOR
1:51
How does Emot explain diffusion models versus the standard transformers?
speaker_0NARRATOR
1:56
He contrasts autoregressive transformers with diffusion models by explaining that diffusion works by destroying data with noise, then learning to reconstruct it.
speaker_0NARRATOR
2:06
He sees that process as tied to compression, least action, and even how nature itself seems to minimize effort while preserving structure.
speaker_1NARRATOR
2:15
And stable diffusion itself was meant to democratize this technology, right? Exactly.
speaker_1ADVERTISER
12:59
Fresh for everyone.
Dagogo AltraideHOST
13:01
DALI2 actually generates the image using a process called diffusion, which has been described by the company as starting with a bag of dots and then filling in a pattern with greater and greater detail.
Dagogo AltraideHOST
13:12
Diffusion is the hottest new method of AI generation and can be a whole other topic for another day.
Dagogo AltraideHOST
13:17
But in order to save time, let's move on to something that I find a lot more interesting.
Dagogo AltraideHOST
13:21
The mimicking of human preference.

Unknown podcast

An Empirical Study of Training Pixel-Space Text-to-Image Diffusion Models

Aug 19 · 11 Mentions

AshleyHOST
0:06
Today, we're diving into a paper from the Hugging Face Daily Paper list of August 18, 2026.
AshleyHOST
0:12
It has garnered 26 upvotes and is titled An Empirical Study of Training Pixel Space Text-to-Image Diffusion Models.
speaker_0HOST
0:20
The first two authors are Deng Yangjiang and Ruoyidu, with the corresponding authors being Peng Gao and Harry Yang.
AshleyHOST
0:28
They come from a collaboration between Alibaba Token Hub of Alibaba Group and the Hong Kong University of Science and Technology.
speaker_0HOST
0:36
So, Ashley, let's dive into the introduction.
speaker_0HOST
0:40
What's the main focus of this paper?
AshleyHOST
0:42
The primary focus is on pixel space diffusion models for generative tasks, specifically text-to-image synthesis.
AshleyHOST
0:50
This is an area that has seen significant interest but hasn't been as thoroughly explored as its latent space counterpart.

We value your privacy

We use cookies to understand how you use our platform and to improve your experience. Click “Accept All” to consent, or “Decline non-essential” to opt out of non-essential cookies. Read our Privacy Policy.