
Autoencoder
1
MENTIONS
1
EPISODES
1
PODCASTS
Search complete. 1 mention across 1 episode found for "Autoencoder".
Sep 24, 2026
“WorkspaceBench: Evaluating Interpretability Methods for the Global Workspace” by camilablank, agam_bhatia, Euan Ong, Neel Nanda
T
6:43Type Three AudioNARRATOR
NLAs are not trained to target the workspace specifically, so their text may describe any content recoverable from the activation.
T
6:51Type Three AudioNARRATOR
Sparse Autoencoders decomposes an activation into a sparse combination of learned dictionary features.
T
6:59Type Three AudioNARRATOR
An activation is read out as the descriptions of its most strongly active features.
T
7:04Type Three AudioNARRATOR
Like NLAs, SAEs are not trained to target the workspace specifically and readout quality depends on how accurate the feature labels are.