
Theo and Ben break down a packed week of AI releases, from Opus 5.5's insane release to Grok 4.7 falling short of the hype. They then dig into why post-training matters, where Jev, GPT-6 Sol, and Luna fit, and more!Thank you to General Translation & Paper for Sponsoring!General Translation: https://nerdsnipe.link/gtPaper: https://nerdsnipe.link/paperSources available on our Substack:https://nerdsnipe.substack.com/Listen wherever you get your podcasts:Spotify: https://nerdsnipe.link/spotifyApple: https://nerdsnipe.link/appleElsewhere: https://nerdsnipe.link/listenTimestamps:0:00 Intro3:18 GPT-6 Sol & Luna14:58 Grok 4.738:47 Jev explained52:48 Why model routing fails1:05:06 Opus 5.51:26:17 Pricing & usage limits1:34:23 Building games with Opus1:41:10 Porting TypeScript to Rust1:48:57 Agent workflows & testing
Sep 30
2 hr 8 min

Theo & Ben break down Jacob Coxon's resignation after three years in pretraining at OpenAI and Anthropic—and his warning that both labs are racing toward AGI and gambling with our lives—along with Dario's latest blog post about pacing the frontier and how Apple's iPhone Duo could have ruined app development (if AI didn't exist).Thanks to this episode's sponsor, PostHog:PostHog: https://nerdsnipe.link/posthogListen wherever you get your podcasts:Spotify: https://nerdsnipe.link/spotifyApple: https://nerdsnipe.link/appleElsewhere: https://nerdsnipe.link/listenSources available on our Substack:https://nerdsnipe.substack.com/Timestamps:0:00 Intro2:53 New iPhones22:24 Apple and Carriers32:38 Astra’s Reliability47:55 Codex Usage and Caching1:19:44 Anthropic Resignation1:38:38 Pacing the Frontier1:40:02 Embedded AI Evaluators1:43:10 The Hugging Face Hack2:02:29 React Native Debate2:09:18 Fast Inference Tradeoffs2:12:40 Democratic AI Coordination2:18:22 China and Global Pacing
Sep 19
2 hr 32 min

Theo & Ben break down their latest thoughts on how Astra and Fable 5.1's releases have changed the way they work, how they interact in T3 Code, and then circle back to the latest with Codex usage limits, Gemini 3.8 Flash, Muse Spark 1.3, and the growing gap between model benchmarks and real coding workflows.Thanks to this episode's sponsor, General Translation: General Translation: https://nerdsnipe.link/gtListen wherever you get your podcasts: Spotify: https://nerdsnipe.link/spotify Apple: https://nerdsnipe.link/apple Elsewhere: https://nerdsnipe.link/listenSources available on our Substack: https://nerdsnipe.substack.com/Timestamps:00:00 Intro04:57 Gemini 3.8 Flash14:48 Muse Spark & small models22:08 AI subscriptions34:14 GPT-6 Astra launch & pricing56:04 Rate limits & autonomous coding1:10:14 Fable 5.11:23:23 AI design & 3D demos1:40:53 Astra vs. Fable: Which to use?
Sep 10
2 hr 5 min

Theo & Ben break down OpenAI's latest model, Astra, and why it's their new benchmark for AI coding, computer use, and multimodal work. But it's not all good: they explain why its UI, stopping behavior, and agent reliability still create friction in real software workflows. From DEFCON puzzles to coding-agent PRs, we're comparing Astra with Fable and asking what a trustworthy OpenAI model should do next.Thank you to PostHog for sponsoring today's episode!PostHog, all-in-one suite of product tools: https://nerdsnipe.link/posthogListen wherever you get your podcasts:Spotify: https://nerdsnipe.link/spotifyApple: https://nerdsnipe.link/appleElsewhere: https://nerdsnipe.link/listenSources available on our Substack:https://nerdsnipe.substack.com/Timestamps00:00 Meet Astra04:03 Best Model Ever, With Catches06:33 Reasoning and 3D Benchmarks20:00 Astra Rebuilds Ping.gg30:04 Instruction-Following Problems44:47 The Uncommitted Fix Debate01:09:37 The PR Babysitting Failure01:22:41 Why Fable Still Wins01:30:41 Multimodal and Computer Use01:39:29 Astra vs. Fable Fleet Data
Sep 3
1 hr 51 min

Theo & Ben break down OpenAI's GPT-5.6 Sol price cut, breakdown the "Ox Alpha" stealth model we now know is GLM 5.3 Flash, then take a drink every time they say "Grok" while ranking every current AI model on a tier list! What could go wrong?Thanks to this episode's sponsor, General Translation:General Translation: https://nerdsnipe.link/gtListen wherever you get your podcasts:Spotify: https://nerdsnipe.link/spotifyApple: https://nerdsnipe.link/appleElsewhere: https://nerdsnipe.link/listenSources available on our Substack:https://nerdsnipe.substack.com/Timestamps0:00 Intro3:30 OpenAI vs. Anthropic14:44 Anthropic’s delayed models20:53 Kimi K3 and open weights30:08 The Alpha stealth model48:26 Model tier list begins1:09:53 GPT-5.6 Luna1:13:09 Opus and Sonnet1:20:25 Gemini models1:31:15 DeepSeek and local models1:59:58 Fable vs. Sol2:28:46 Final rankings
Aug 28
2 hr 42 min

Anthropic started watermarking their model's outputs, Dario posted on X, SpaceXAi is getting really good really fast, Meta is shipping again, and we won DEF CON 2026. Thank you PostHog, the all in one suite of product tools for sponsoring today's Episode! Check them out at: nerdsnipe.link/posthogListen wherever you get your podcasts: - Spotify: nerdsnipe.link/spotify- Apple: nerdsnipe.link/apple- Elsewhere: nerdsnipe.link/listenSources available on our Substack: nerdsnipe.substack.comTimestamps:0:00 Intro2:56 Claude Watermarks15:04 Meta + Muse48:48 GLM-5.355:15 Qwen + DeepSeek1:08:48 Grok 4.6 + Bot1:20:15 Gavin vs Dario1:43:44 DEFCON2:15:35 Viewer Q&A
Aug 20
2 hr 18 min

Theo and Ben break down why Opus 5 feels like GPT-5.5 crossed with Fable rather than GPT-5.6 crossed with Fable, what Fable found when it audited an Opus thread line by line, where Opus still clearly wins (3D, animation, and Claude Code limits at half Fable's price with no 50% weekly cap), plus Kimi K3, GLM 5.2, the Hugging Face hack, Grok 4.5, Codex, and T3 Code.Thanks to this episode's sponsor, General Translation:General Translation: https://nerdsnipe.link/gtSources available on our Substack:https://nerdsnipe.substack.com/
Jul 28
1 hr 23 min

This week Grok 4.5, Muse Spark 1.1, GPT-Live, and GPT 5.6 all dropped. Theo and Ben break down which models you should care about, how OpenAI fumbled the Codex to ChatGPT app transition, the abysmal usage numbers for 5.6, and the latest drama surrounding OpenAI and Sam Altman. Plus, how many satellites would it take to trap humanity on Earth, and how many data centers to heat up the ocean?Thank you to this episode's sponsors: Composio & WorkOs.https://nerdsnipe.link/composiohttps://nerdsnipe.link/workosSources available on Substack: https://nerdsnipe.substack.com/
Jul 14
2 hr 6 min

We've spent six figures in tokens testing OpenAI's 5.6 Sol model to see whether if its better than Fable and what OpenAI have done to make it even better than 5.5. Also, we breakdown why we both moved our agents to Linux boxes, how to actually burn $65k on a single loop, and the Codex vs Claude Code subagent gap that's now bigger than the model gap itself.Thanks to this episode's sponsors Clerk and General Translation:Clerk, the auth platform with the best DX: https://nerdsnipe.link/clerkGeneral Translation: https://nerdsnipe.link/gtSources Available on our Substack:https://nerdsnipe.substack.com/
Jul 9
1 hr 11 min

Fable was supposed to give us 14 days and instead we got 3 before the export ban. Now it's back at half the rate limits. Also, we break down why AI-generated Claude Code skills fall apart, what actually triggers Fable-to-Opus rerouting, the Mythos export-control timeline, and whether GLM 5.2 can really touch Sonnet 5.Thank you to PostHog for sponsoring today's episode! PostHog, all in one suite of product tools: https://nerdsnipe.link/posthogSources available on our Substack:https://nerdsnipe.substack.com/
Jul 8
1 hr 16 min
Load more
