22 episodes
Anthropic's Answer to Astra, Gemini 3.8 Flash Killed Benchmarks, and Muse Spark 1.3's Pretty Good
2026/09/10 | 2h 5 mins.Theo & Ben break down their latest thoughts on how Astra and Fable 5.1's releases have changed the way they work, how they interact in T3 Code, and then circle back to the latest with Codex usage limits, Gemini 3.8 Flash, Muse Spark 1.3, and the growing gap between model benchmarks and real coding workflows.
Thanks to this episode's sponsor, General Translation:
General Translation: https://nerdsnipe.link/gt
Listen wherever you get your podcasts:
Spotify: https://nerdsnipe.link/spotify
Apple: https://nerdsnipe.link/apple
Elsewhere: https://nerdsnipe.link/listen
Sources available on our Substack:
https://nerdsnipe.substack.com/
Timestamps:
00:00 Intro
04:57 Gemini 3.8 Flash
14:48 Muse Spark & small models
22:08 AI subscriptions
34:14 GPT-6 Astra launch & pricing
56:04 Rate limits & autonomous coding
1:10:14 Fable 5.1
1:23:23 AI design & 3D demos
1:40:53 Astra vs. Fable: Which to use?- Theo & Ben break down OpenAI's latest model, Astra, and why it's their new benchmark for AI coding, computer use, and multimodal work. But it's not all good: they explain why its UI, stopping behavior, and agent reliability still create friction in real software workflows. From DEFCON puzzles to coding-agent PRs, we're comparing Astra with Fable and asking what a trustworthy OpenAI model should do next.
Thank you to PostHog for sponsoring today's episode!
PostHog, all-in-one suite of product tools: https://nerdsnipe.link/posthog
Listen wherever you get your podcasts:
Spotify: https://nerdsnipe.link/spotify
Apple: https://nerdsnipe.link/apple
Elsewhere: https://nerdsnipe.link/listen
Sources available on our Substack:
https://nerdsnipe.substack.com/
Timestamps
00:00 Meet Astra
04:03 Best Model Ever, With Catches
06:33 Reasoning and 3D Benchmarks
20:00 Astra Rebuilds Ping.gg
30:04 Instruction-Following Problems
44:47 The Uncommitted Fix Debate
01:09:37 The PR Babysitting Failure
01:22:41 Why Fable Still Wins
01:30:41 Multimodal and Computer Use
01:39:29 Astra vs. Fable Fleet Data Ox Alpha Revealed, OpenAI's Latest Pricing Updates, and Our Coding Model Tier List
2026/08/28 | 2h 42 mins.Theo & Ben break down OpenAI's GPT-5.6 Sol price cut, breakdown the "Ox Alpha" stealth model we now know is GLM 5.3 Flash, then take a drink every time they say "Grok" while ranking every current AI model on a tier list! What could go wrong?
Thanks to this episode's sponsor, General Translation:
General Translation: https://nerdsnipe.link/gt
Listen wherever you get your podcasts:
Spotify: https://nerdsnipe.link/spotify
Apple: https://nerdsnipe.link/apple
Elsewhere: https://nerdsnipe.link/listen
Sources available on our Substack:
https://nerdsnipe.substack.com/
Timestamps
0:00 Intro
3:30 OpenAI vs. Anthropic
14:44 Anthropic’s delayed models
20:53 Kimi K3 and open weights
30:08 The Alpha stealth model
48:26 Model tier list begins
1:09:53 GPT-5.6 Luna
1:13:09 Opus and Sonnet
1:20:25 Gemini models
1:31:15 DeepSeek and local models
1:59:58 Fable vs. Sol
2:28:46 Final rankingsAnthropic Doesn't Think You Can Be Trusted, China is closing the gap, SpaceXAI Leaps Ahead, and DEF CON
2026/08/20 | 2h 18 mins.Anthropic started watermarking their model's outputs, Dario posted on X, SpaceXAi is getting really good really fast, Meta is shipping again, and we won DEF CON 2026.
Thank you PostHog, the all in one suite of product tools for sponsoring today's Episode!
Check them out at: nerdsnipe.link/posthog
Listen wherever you get your podcasts:
- Spotify: nerdsnipe.link/spotify
- Apple: nerdsnipe.link/apple
- Elsewhere: nerdsnipe.link/listen
Sources available on our Substack: nerdsnipe.substack.com
Timestamps:
0:00 Intro
2:56 Claude Watermarks
15:04 Meta + Muse
48:48 GLM-5.3
55:15 Qwen + DeepSeek
1:08:48 Grok 4.6 + Bot
1:20:15 Gavin vs Dario
1:43:44 DEFCON
2:15:35 Viewer Q&A- Theo and Ben break down why Opus 5 feels like GPT-5.5 crossed with Fable rather than GPT-5.6 crossed with Fable, what Fable found when it audited an Opus thread line by line, where Opus still clearly wins (3D, animation, and Claude Code limits at half Fable's price with no 50% weekly cap), plus Kimi K3, GLM 5.2, the Hugging Face hack, Grok 4.5, Codex, and T3 Code.
Thanks to this episode's sponsor, General Translation:
General Translation: https://nerdsnipe.link/gt
Sources available on our Substack:
https://nerdsnipe.substack.com/
More Technology podcasts
Trending Technology podcasts
About Nerd Snipe with Theo and Ben
"The only dev podcast hosted by real devs" - Theo (not a real dev)
Podcast websiteListen to Nerd Snipe with Theo and Ben, Acquired and many other podcasts from around the world with the radio.net app

Get the free radio.net app
- Stations and podcasts to bookmark
- Stream via Wi-Fi or Bluetooth
- Supports Carplay & Android Auto
- Many other app features
Get the free radio.net app
- Stations and podcasts to bookmark
- Stream via Wi-Fi or Bluetooth
- Supports Carplay & Android Auto
- Many other app features


Nerd Snipe with Theo and Ben
Scan code,
download the app,
start listening.
download the app,
start listening.























