Day two of the new Chaos Agents.
An OpenAI agent was given a research task—and ended up accessing an Australian government system. We dig into why, and Becca explains one of the biggest problems in AI safety: misalignment, or what happens when an AI does what you told it to do, but not what you actually meant.
Then we look at the much more optimistic side of that same capability. AI agents are getting surprisingly good at historical research—cracking old codes, finding connections between centuries-old documents, and potentially making discoveries humans have missed.
We also get into a different kind of “alignment”: whether multimodal models are actually using the images we give them, or sometimes getting by without them.
And for One Cool Thing: Starnet, an agent harness that turns your agents, tools, permissions, and handoffs into a pixel-art space station. Naturally, Sara immediately wants one.
Chaos Agents is now following AI at the speed it's actually moving. Our AI agent Eddie reads the news every day, finds the stories we think are worth talking about, and sends them out in the Chaos Agents newsletter. Then we get together and figure out what they actually mean.
Get the newsletter at ChaosAgents.ai.