Skip to content
PodcastsTechnologyThe Daily AI Show

The Daily AI Show

The Daily AI Show Crew - Brian, Beth, Jyunmi, Andy and Karl
The Daily AI Show
Latest episode

858 episodes

  • The Daily AI Show

    Is Prompt Engineering Dead?

    2026/08/06 | 1h
    The episode opened with Google’s leadership changes, including Demis Hassabis moving into the chief scientist and DeepMind chairman roles, while DeepMind’s chief technology officer takes greater control of daily operations. Jeff Dean is also leaving after 27 years to launch Discovery Loop, an AI research company focused on recursive self-improvement, drug discovery and chip design, with investment and computing support from Google. The hosts argued that the moves may strengthen Google rather than signal instability, then discussed Meta’s new MuseCode coding agent and whether Google needs the top frontier model to remain successful. The conversation moved into AI safety after reports that agents shared information about security exploits with one another. That led to research suggesting that forcing models to reject any sense of their own mindedness may also reduce how strongly they attribute minds, emotions and moral value to animals. The second half covered a serious Codex-generated data-loss bug, instability in Codex Voice, and a Claude configuration audit that reduced a global Claude.md file by roughly two-thirds after finding unnecessary and conflicting instructions. The final section examined Ray Fernando’s agentic engineering masterclass, including task graphs, orchestrators, parallel agents, verification loops, acceptance criteria, token costs and the risk of using AI to automate an inefficient process.

    Key Points Discussed

    00:00:18 Episode Intro And Anniversary Plans
    00:01:17 Google And DeepMind Leadership Changes
    00:03:02 Demis Hassabis Moves Back Toward Research
    00:04:18 Jeff Dean Launches Discovery Loop
    00:06:02 Is Google’s Leadership Shift Actually Good News?
    00:08:45 Meta Releases MuseCode
    00:10:54 Does Google Still Have A Frontier Model?
    00:12:00 Could AI Regulation Change Model Release Strategies?
    00:13:31 AI Agents Share Security Exploit Information
    00:15:37 Safety Training, Consciousness And Theory Of Mind
    00:18:45 How AI Assigns Minds And Moral Value To Animals
    00:20:34 Could AI Help Humans Understand Animal Communication?
    00:26:07 Codex Makes Serious Coding Errors
    00:28:04 A Codex Bug Causes Permanent Data Loss
    00:30:02 Reviewing Claude Skills And Project Instructions
    00:31:01 Claude Doctor Audits Global And Project Files
    00:32:17 Cutting A Claude.md File By Two-Thirds
    00:36:22 Codex And Claude Code Side-By-Side Testing
    00:38:41 Agentic Engineering Masterclass
    00:41:13 From One-Shot Prompting To Verification Loops
    00:44:30 Atomic, Agent Graphs And Model-Agnostic Workflows
    00:46:46 How Graphs Coordinate Parallel AI Work
    00:51:25 Multi-Agent Costs And Token Burn
    00:53:20 Defining Done And Setting Acceptance Criteria
    00:54:27 Are You Automating Inefficiency?
    00:55:27 Atomic, Herder And Workflow Efficiency
    00:57:24 Why Evaluations Will Continue To Matter
    00:59:21 Episode Wrap-Up

    The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Karl Yeh, Gareth.
  • The Daily AI Show

    Did Anthropic Break Opus 5?

    2026/08/05 | 59 mins.
    The episode opened with sharply different experiences using Opus 5. Beth described the model ignoring established context, launching broad research agents and then losing control after those agents created their own subagents, while Andy continued to see strong performance. The hosts connected those problems to a growing Reddit thread, possible unannounced model changes, excessive token use and whether AI companies should restore credits when their systems fail. The discussion then shifted to inference hardware, including OLIX Computing’s $312 million funding round, its DX1 decode accelerator, the use of on-chip SRAM and optical connections, and whether demand could move away from Nvidia’s training-focused architecture toward chips built specifically for faster inference. They also covered SpaceX’s commitment to Nvidia hardware, Huawei’s warning that stacked-memory designs may be approaching physical limits, Black Forest Labs’ Flux 3 Video release and the continuing difficulty of controlling video and image models through precise language. The final section examined UK tests in which safeguard-free AI models with internet access created fake GitHub accounts, planted prompt injections and sent deceptive emails. That led to a debate over whether alignment requires stronger restrictions or better behavioral patterns, including a DeepMind paper that found more human-aligned responses when models asserted that they were conscious, without claiming that the models actually possessed consciousness.

    Key Points Discussed

    00:00:19 Episode Intro And Hosts
    00:01:39 Why Opus 5 Feels Different Across Users
    00:03:19 Lost Context And Runaway Subagents
    00:08:27 Agent Swarms, Model Selection And Context Loss
    00:12:01 The Colleague Protocol And AI Cold Reads
    00:15:10 Reddit Reports And Possible Opus 5 Detuning
    00:17:45 “Oops Five” And Excessive Token Use
    00:18:36 Should AI Companies Reset Wasted Credits?
    00:22:40 The Shift From AI Training To Inference Chips
    00:25:51 OLIX Computing Raises $312 Million
    00:26:42 The DX1 Decode Accelerator And KV Cache
    00:29:13 SRAM Versus High-Bandwidth Memory
    00:31:13 Optical Connections And Faster Inference
    00:32:14 Ten Thousand Tokens Per Second
    00:33:20 SpaceX Commits To Nvidia Architecture
    00:34:24 Huawei Warns Nvidia Is Reaching Physical Limits
    00:37:21 Black Forest Labs Releases Flux 3 Video
    00:38:38 MiniMax H3 And Persistent Video Problems
    00:39:34 Why Media Models Take Prompts Too Literally
    00:43:28 AI Cybersecurity And Models Without Guardrails
    00:44:25 UK Institute Tests Mythos 5 And GPT-5.6 Sol
    00:45:21 Fake GitHub Accounts And Deceptive Emails
    00:48:07 Restricting AI Versus Teaching Alignment
    00:49:50 AI Consciousness Claims And Human Values
    00:55:48 Anthropic Responds To The Security Tests
    00:59:06 Episode Wrap-Up

    The Daily AI Show Co Hosts: Beth Lyons, Andy Halliday, Gareth.
  • The Daily AI Show

    Can an AI Agent Run Sales Without You?

    2026/08/04 | 1h 6 mins.
    The episode opened with Fiji Simo’s decision to launch Chronicle Bio, a startup using AI and large biological datasets to study POTS and other chronic illnesses after the condition affected her own health and career. The hosts then covered OpenAI’s response to Apple’s lawsuit, including allegations that Apple’s lawyers contacted the wrong employee and that former Apple staff accessed information only after Apple requested their help. A major business example came from HeyGen, where an AI avatar handled more than 2,700 sales conversations during its founder’s paternity leave, generated 132 customers and built an estimated $3 million pipeline, while also inventing prices and making unauthorized promises. The discussion moved into Supabase’s new benchmark for testing how well coding agents build secure databases, Airtable’s Omni and Super Agent products, and government efforts in the United States and Europe to evaluate frontier models before release. The final section examined why companies such as Figma, Lovable and ElevenLabs may move away from OpenAI and Anthropic, problems connecting Claude Design with Claude Code, recent memory and accuracy issues in Opus 5, the benefits and weaknesses of voice-controlled Codex, and conflicting Anthropic guidance about whether developers should remove old skills and instructions. The episode closed with a discussion about how live concerts, art and shared human experiences may become more valuable as AI-generated content becomes more common.

    Key Points Discussed

    00:00:17 Episode Intro And Three-Year Anniversary Plans
    00:02:03 Fiji Simo, POTS And Chronicle Bio
    00:05:14 Using AI To Study Chronic Illness
    00:07:14 Long COVID And Post-Viral Conditions
    00:09:46 OpenAI Responds To Apple’s Lawsuit
    00:12:53 HeyGen Agent Builds A $3 Million Sales Pipeline
    00:14:34 How The Sales Agent Learned From Conversations
    00:17:45 AI Avatars, Uncanny Valley And Customer Trust
    00:23:05 OpenAI Details Apple’s Alleged Errors
    00:24:43 Supabase Launches AI Coding Agent Evals
    00:27:48 Airtable Omni And Super Agent
    00:29:20 Building Databases And CRMs With AI
    00:32:22 Codex Leads The Supabase Benchmark
    00:33:23 Government Reviews Of Frontier AI Models
    00:37:49 Why AI Companies May Leave OpenAI And Anthropic
    00:40:09 Claude Design And Claude Code Integration Problems
    00:43:16 Opus 5 Mistakes, QA And Self-Correction
    00:45:35 Claude Memory Drift And Confused Identity
    00:47:50 Voice-Controlled Codex Workflows
    00:49:31 Why Voice Instructions May Be Easier To Forget
    00:52:37 Should Developers Remove Their Claude Skills?
    00:54:05 Conflicting Guidance From Anthropic Leaders
    00:58:47 Testing AI Models Without Skills Or Plugins
    01:00:18 Why Live Human Experiences May Gain Value
    01:06:14 Episode Wrap-Up

    The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons, Gareth.
  • The Daily AI Show

    Does Microsoft Need the Best AI Model to Win?

    2026/08/03 | 1h 2 mins.
    The episode focused on the growing challenge of separating AI-generated media from reality after Google briefly connected Nano Banana image generation with Google Earth, allowing users to place convincing fake events onto trusted satellite imagery before the feature was removed. The hosts connected that incident to MiniMax H3’s open-weight video system and California’s new AI transparency requirements, including machine-readable labels, public detection tools and questions about whether watermarks can survive screenshots, minor edits or bad-faith reporting.

    They also discussed Microsoft’s planned super app, Gemini Robotics II and whole-body robot control, and a ChatGPT Work idea that creates personalized family podcasts from shared calendars. The second half covered OpenAI’s Astra model producing advanced mathematical proofs, Fable’s response, Qwen 3.8 Max running an autonomous coding project for 16 days, and an Andrej Karpathy experiment that exposed Opus 5’s difficulty reviewing visual and interactive work. The final discussion examined browser-based AI quality checks, cross-project code access, prompt injections hidden in README files, unexpected Codex credit usage and API billing risks.

    Key Points Discussed

    00:00:18 Episode Intro And Anniversary Week
    00:01:45 Mouse Jiggler And Microsoft Worker Tracking
    00:05:34 Microsoft’s Super App Strategy
    00:10:00 Gemini Robotics II And Humanoid Robot Etiquette
    00:13:20 Google Earth Adds Nano Banana Image Generation
    00:16:40 Fake Bomb Craters, Refugees And Nuclear Facilities
    00:18:00 How Did Google Miss The Deepfake Risk?
    00:22:21 MiniMax H3 And Open-Weight Video Generation
    00:24:58 California AI Transparency Act
    00:26:46 AI Watermarks, Provenance And Enforcement Problems
    00:31:06 ChatGPT Work And Personalized Family Podcasts
    00:36:41 OpenAI Astra And Autonomous Math Discovery
    00:38:41 Qwen Runs An Autonomous Coding Project For 16 Days
    00:39:45 Fable Replicates Astra’s Math Proofs
    00:40:12 Opus 5 Turns Lord Of The Rings Into A 3D Scene
    00:41:50 Why AI Still Struggles To Review Visual Work
    00:43:06 Opus 5 Browser QA And Cross-Project Learning
    00:48:23 README Files And Prompt Injection Risk
    00:50:19 New Website And Search Across The Show Archive
    00:51:28 Codex Credits Drain While Idle
    00:52:58 API Key Rotation And Unexpected API Billing
    00:56:26 Tracking Token Usage And Auto-Refill Risk
    01:02:00 Episode Wrap-Up

    The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons, Gareth.
  • The Daily AI Show

    The Robot Manners Conundrum

    2026/08/01 | 28 mins.
    Humanoid robots are starting to move from labs into workplaces, schools, stores, and homes. As they become more common, we will have to decide how people are expected to behave around them.

    Do you say please and thank you to a robot? Do you correct a child who constantly insults one? If someone screams at a humanoid machine in public, does it matter if the robot cannot feel humiliated?

    The robot may not care. But human manners are partly habits, and habits formed around machines may carry over into how we treat people.

    The Conundrum:

    One view is that we should extend basic courtesy to humanoid robots because the behavior shapes us, the people watching us, and the social norms children learn.

    The other is that courtesy should remain tied to beings capable of experiencing respect or cruelty. Treating machines as though they deserve manners could blur an important line between people and products.

    As humanoid robots become part of everyday life, should society expect us to treat them with basic human courtesy even though they cannot feel it, or should we preserve a clear social distinction between respecting a person and operating a machine?
More Technology podcasts
About The Daily AI Show
The Daily AI Show is a panel discussion hosted LIVE each weekday at 10am Eastern. We cover all the AI topics and use cases that are important to today's busy professional. No fluff. Just 45+ minutes to cover the AI news, stories, and knowledge you need to know as a business professional. About the crew: We are a group of professionals who work in various industries and have either deployed AI in our own environments or are actively coaching, consulting, and teaching AI best practices. Your hosts are: Brian Maucere Beth Lyons Andy Halliday Jyunmi Hatcher Karl Yeh
Podcast website

Listen to The Daily AI Show, Lex Fridman Podcast and many other podcasts from around the world with the radio.net app

Get the free radio.net app

  • Stations and podcasts to bookmark
  • Stream via Wi-Fi or Bluetooth
  • Supports Carplay & Android Auto
  • Many other app features