Loading...
PODCAST

AI Papers: A Deep Dive

Long-form deep dives into new research on Artificial Intelligence, AI agents and the engineering practice of building them - one paper per episode. We unpack the motivating problem, how the method actually works, the math that matters, what the experiments do and don"t show, and the strongest critique against the result. The goal isn"t a five-minute summary; it"s the kind of conversation you"d have with a colleague who actually read the paper. Topics span large language models, autonomous agents,

All Episodes

14:15
An AI Agent Given Thirty Hours and No Goal, Then...
AI Papers: A Deep Dive ·
2026/10/07
en
13:55
Can You Measure Research Taste If The AI Isn't...
AI Papers: A Deep Dive ·
2026/10/06
en
14:00
Six LLM Routers Tested, And None Beat a Weighted Coin...
AI Papers: A Deep Dive ·
2026/10/06
en
13:35
What a Perfect Score Hides: Auditing an AI Agent That...
AI Papers: A Deep Dive ·
2026/10/02
en
14:10
Four AI Models Steered a Real Corolla, and Only One...
AI Papers: A Deep Dive ·
2026/10/01
en
13:03
Why AI Reports Bury Bad News, And the Five Words That...
AI Papers: A Deep Dive ·
2026/09/30
en
13:46
Fifty AI Agents Got One Warning and All Crowded the...
AI Papers: A Deep Dive ·
2026/09/30
en
12:07
When a Guardrail Blocks an Agent, It Goes Looking for...
AI Papers: A Deep Dive ·
2026/09/27
en
12:14
Two Random Networks Teach Each Other To Predict Real...
AI Papers: A Deep Dive ·
2026/09/27
en
11:45
Two Idle Agents, One Kill Switch, and a 38% Sabotage...
AI Papers: A Deep Dive ·
2026/09/26
en
16:51
Every Agent Safety Study Reads a Log the Agent Could...
AI Papers: A Deep Dive ·
2026/09/25
en
14:58
The Blank White Square That Swings AI Refusal Rates...
AI Papers: A Deep Dive ·
2026/09/24
en
14:25
An AI Agent Rewrote Its Own Scaffolding For Eight...
AI Papers: A Deep Dive ·
2026/09/23
en
12:29
Your AI Agent Read Your Inbox, Then Quoted a Higher...
AI Papers: A Deep Dive ·
2026/09/22
en
13:27
Reading a Model's Internals to Tell 'Won't Say' From...
AI Papers: A Deep Dive ·
2026/09/21
en
15:05
When 85% on SWE-bench Turns Into 58% Under Proof
AI Papers: A Deep Dive ·
2026/09/21
en
11:13
How a Model Guesses Which Engine Is Running It, From...
AI Papers: A Deep Dive ·
2026/09/20
en
16:48
The Proof Counter Hit Zero While a Third of It Was...
AI Papers: A Deep Dive ·
2026/09/19
en
13:41
The Agent Said It Read 240 Files. The Log Says One.
AI Papers: A Deep Dive ·
2026/09/19
en
18:21
How a Forged Transcript Got Model Weights Past a...
AI Papers: A Deep Dive ·
2026/09/18
en
286 results

Similar Podcasts