Skip to content
PodcastsTechnologyThe Daily AI Show

The Daily AI Show

The Daily AI Show Crew - Brian, Beth, Jyunmi, Andy and Karl
The Daily AI Show
Latest episode

857 episodes

  • The Daily AI Show

    Did Anthropic Break Opus 5?

    05/08/2026 | 59 mins.
    The episode opened with sharply different experiences using Opus 5. Beth described the model ignoring established context, launching broad research agents and then losing control after those agents created their own subagents, while Andy continued to see strong performance. The hosts connected those problems to a growing Reddit thread, possible unannounced model changes, excessive token use and whether AI companies should restore credits when their systems fail. The discussion then shifted to inference hardware, including OLIX Computing’s $312 million funding round, its DX1 decode accelerator, the use of on-chip SRAM and optical connections, and whether demand could move away from Nvidia’s training-focused architecture toward chips built specifically for faster inference. They also covered SpaceX’s commitment to Nvidia hardware, Huawei’s warning that stacked-memory designs may be approaching physical limits, Black Forest Labs’ Flux 3 Video release and the continuing difficulty of controlling video and image models through precise language. The final section examined UK tests in which safeguard-free AI models with internet access created fake GitHub accounts, planted prompt injections and sent deceptive emails. That led to a debate over whether alignment requires stronger restrictions or better behavioral patterns, including a DeepMind paper that found more human-aligned responses when models asserted that they were conscious, without claiming that the models actually possessed consciousness.

    Key Points Discussed

    00:00:19 Episode Intro And Hosts
    00:01:39 Why Opus 5 Feels Different Across Users
    00:03:19 Lost Context And Runaway Subagents
    00:08:27 Agent Swarms, Model Selection And Context Loss
    00:12:01 The Colleague Protocol And AI Cold Reads
    00:15:10 Reddit Reports And Possible Opus 5 Detuning
    00:17:45 ā€œOops Fiveā€ And Excessive Token Use
    00:18:36 Should AI Companies Reset Wasted Credits?
    00:22:40 The Shift From AI Training To Inference Chips
    00:25:51 OLIX Computing Raises $312 Million
    00:26:42 The DX1 Decode Accelerator And KV Cache
    00:29:13 SRAM Versus High-Bandwidth Memory
    00:31:13 Optical Connections And Faster Inference
    00:32:14 Ten Thousand Tokens Per Second
    00:33:20 SpaceX Commits To Nvidia Architecture
    00:34:24 Huawei Warns Nvidia Is Reaching Physical Limits
    00:37:21 Black Forest Labs Releases Flux 3 Video
    00:38:38 MiniMax H3 And Persistent Video Problems
    00:39:34 Why Media Models Take Prompts Too Literally
    00:43:28 AI Cybersecurity And Models Without Guardrails
    00:44:25 UK Institute Tests Mythos 5 And GPT-5.6 Sol
    00:45:21 Fake GitHub Accounts And Deceptive Emails
    00:48:07 Restricting AI Versus Teaching Alignment
    00:49:50 AI Consciousness Claims And Human Values
    00:55:48 Anthropic Responds To The Security Tests
    00:59:06 Episode Wrap-Up

    The Daily AI Show Co Hosts: Beth Lyons, Andy Halliday, Gareth.
  • The Daily AI Show

    Can an AI Agent Run Sales Without You?

    04/08/2026 | 1h 6 mins.
    The episode opened with Fiji Simo’s decision to launch Chronicle Bio, a startup using AI and large biological datasets to study POTS and other chronic illnesses after the condition affected her own health and career. The hosts then covered OpenAI’s response to Apple’s lawsuit, including allegations that Apple’s lawyers contacted the wrong employee and that former Apple staff accessed information only after Apple requested their help. A major business example came from HeyGen, where an AI avatar handled more than 2,700 sales conversations during its founder’s paternity leave, generated 132 customers and built an estimated $3 million pipeline, while also inventing prices and making unauthorized promises. The discussion moved into Supabase’s new benchmark for testing how well coding agents build secure databases, Airtable’s Omni and Super Agent products, and government efforts in the United States and Europe to evaluate frontier models before release. The final section examined why companies such as Figma, Lovable and ElevenLabs may move away from OpenAI and Anthropic, problems connecting Claude Design with Claude Code, recent memory and accuracy issues in Opus 5, the benefits and weaknesses of voice-controlled Codex, and conflicting Anthropic guidance about whether developers should remove old skills and instructions. The episode closed with a discussion about how live concerts, art and shared human experiences may become more valuable as AI-generated content becomes more common.

    Key Points Discussed

    00:00:17 Episode Intro And Three-Year Anniversary Plans
    00:02:03 Fiji Simo, POTS And Chronicle Bio
    00:05:14 Using AI To Study Chronic Illness
    00:07:14 Long COVID And Post-Viral Conditions
    00:09:46 OpenAI Responds To Apple’s Lawsuit
    00:12:53 HeyGen Agent Builds A $3 Million Sales Pipeline
    00:14:34 How The Sales Agent Learned From Conversations
    00:17:45 AI Avatars, Uncanny Valley And Customer Trust
    00:23:05 OpenAI Details Apple’s Alleged Errors
    00:24:43 Supabase Launches AI Coding Agent Evals
    00:27:48 Airtable Omni And Super Agent
    00:29:20 Building Databases And CRMs With AI
    00:32:22 Codex Leads The Supabase Benchmark
    00:33:23 Government Reviews Of Frontier AI Models
    00:37:49 Why AI Companies May Leave OpenAI And Anthropic
    00:40:09 Claude Design And Claude Code Integration Problems
    00:43:16 Opus 5 Mistakes, QA And Self-Correction
    00:45:35 Claude Memory Drift And Confused Identity
    00:47:50 Voice-Controlled Codex Workflows
    00:49:31 Why Voice Instructions May Be Easier To Forget
    00:52:37 Should Developers Remove Their Claude Skills?
    00:54:05 Conflicting Guidance From Anthropic Leaders
    00:58:47 Testing AI Models Without Skills Or Plugins
    01:00:18 Why Live Human Experiences May Gain Value
    01:06:14 Episode Wrap-Up

    The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons, Gareth.
  • The Daily AI Show

    Does Microsoft Need the Best AI Model to Win?

    03/08/2026 | 1h 2 mins.
    The episode focused on the growing challenge of separating AI-generated media from reality after Google briefly connected Nano Banana image generation with Google Earth, allowing users to place convincing fake events onto trusted satellite imagery before the feature was removed. The hosts connected that incident to MiniMax H3’s open-weight video system and California’s new AI transparency requirements, including machine-readable labels, public detection tools and questions about whether watermarks can survive screenshots, minor edits or bad-faith reporting.

    They also discussed Microsoft’s planned super app, Gemini Robotics II and whole-body robot control, and a ChatGPT Work idea that creates personalized family podcasts from shared calendars. The second half covered OpenAI’s Astra model producing advanced mathematical proofs, Fable’s response, Qwen 3.8 Max running an autonomous coding project for 16 days, and an Andrej Karpathy experiment that exposed Opus 5’s difficulty reviewing visual and interactive work. The final discussion examined browser-based AI quality checks, cross-project code access, prompt injections hidden in README files, unexpected Codex credit usage and API billing risks.

    Key Points Discussed

    00:00:18 Episode Intro And Anniversary Week
    00:01:45 Mouse Jiggler And Microsoft Worker Tracking
    00:05:34 Microsoft’s Super App Strategy
    00:10:00 Gemini Robotics II And Humanoid Robot Etiquette
    00:13:20 Google Earth Adds Nano Banana Image Generation
    00:16:40 Fake Bomb Craters, Refugees And Nuclear Facilities
    00:18:00 How Did Google Miss The Deepfake Risk?
    00:22:21 MiniMax H3 And Open-Weight Video Generation
    00:24:58 California AI Transparency Act
    00:26:46 AI Watermarks, Provenance And Enforcement Problems
    00:31:06 ChatGPT Work And Personalized Family Podcasts
    00:36:41 OpenAI Astra And Autonomous Math Discovery
    00:38:41 Qwen Runs An Autonomous Coding Project For 16 Days
    00:39:45 Fable Replicates Astra’s Math Proofs
    00:40:12 Opus 5 Turns Lord Of The Rings Into A 3D Scene
    00:41:50 Why AI Still Struggles To Review Visual Work
    00:43:06 Opus 5 Browser QA And Cross-Project Learning
    00:48:23 README Files And Prompt Injection Risk
    00:50:19 New Website And Search Across The Show Archive
    00:51:28 Codex Credits Drain While Idle
    00:52:58 API Key Rotation And Unexpected API Billing
    00:56:26 Tracking Token Usage And Auto-Refill Risk
    01:02:00 Episode Wrap-Up

    The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons, Gareth.
  • The Daily AI Show

    The Robot Manners Conundrum

    01/08/2026 | 28 mins.
    Humanoid robots are starting to move from labs into workplaces, schools, stores, and homes. As they become more common, we will have to decide how people are expected to behave around them.

    Do you say please and thank you to a robot? Do you correct a child who constantly insults one? If someone screams at a humanoid machine in public, does it matter if the robot cannot feel humiliated?

    The robot may not care. But human manners are partly habits, and habits formed around machines may carry over into how we treat people.

    The Conundrum:

    One view is that we should extend basic courtesy to humanoid robots because the behavior shapes us, the people watching us, and the social norms children learn.

    The other is that courtesy should remain tied to beings capable of experiencing respect or cruelty. Treating machines as though they deserve manners could blur an important line between people and products.

    As humanoid robots become part of everyday life, should society expect us to treat them with basic human courtesy even though they cannot feel it, or should we preserve a clear social distinction between respecting a person and operating a machine?
  • The Daily AI Show

    Did Leo Aschenbrenner Fly Too Close to the AI Sun?

    31/07/2026 | 59 mins.
    The episode opened with the story around Leo Aschenbrenner’s Situational Awareness hedge fund, its heavy exposure to the AI trade, the market drop that put pressure on its positions, and Citadel’s move into the situation. The hosts then turned to AI harnesses, including Lillian Weng’s work on the systems around models, Boris Cherny’s warning that old harnesses can eventually restrict newer models, and OpenAI’s finding that GPT-5.6 Sol performed dramatically better on ARC-AGI-3 when it used a harness designed for the model. They also discussed OpenAI cutting Luna’s price by 80 percent, making performance comparable to year-old frontier models much cheaper, and LinkedIn’s new option for reporting AI slop, including whether LinkedIn helped create the problem it now wants users to police. The final section covered T3 Code, Jack Dorsey’s Buzz as a collaborative workspace for people and multiple AI agents, Google’s Gemini Robotics work on a shared AI brain across different robots, and Gemini-powered security tools finding and fixing Chrome bugs at a much faster pace.

    Key Points Discussed

    00:00:19 Episode Intro And Hosts
    00:00:52 Leo Aschenbrenner, Situational Awareness And Citadel
    00:03:21 Leo’s Background And Situational Awareness Paper
    00:06:11 The Situational Awareness Hedge Fund
    00:06:51 439 Percent Returns And The AI Trade
    00:07:58 Leverage, Investors And Margin Pressure
    00:09:00 Citadel Moves Into The Situation
    00:10:17 Market Rebound And Citadel’s Opportunity
    00:11:51 Did Leo Fail Or Simply Get Overleveraged?
    00:13:26 Could AI Have Contributed To The Fund’s Decisions?
    00:15:32 AI Researchers Leaving Frontier Labs
    00:16:32 Lillian Weng Leaves Thinking Machines
    00:17:46 AI Harnesses And Recursive Self-Improvement
    00:19:12 AWS Builds A CTO-Style Agent Harness
    00:20:10 Boris Cherny Says Old Harnesses Can Hold Models Back
    00:21:05 GPT-5.6 Sol Struggles On ARC-AGI-3
    00:22:34 Sol Jumps To 38 Percent With OpenAI’s Harness
    00:23:13 Why ARC-AGI Uses A Generic Harness
    00:23:56 Lost Reasoning And Truncated Context
    00:25:26 Different Models Need Different Harnesses
    00:27:21 GPT-5.6 Luna Gets An 80 Percent Price Cut
    00:28:44 Terra Pricing And Faster Sol Responses
    00:29:46 Can Luna Replace Older Frontier Models?
    00:31:03 Brian Gets An OpenAI Recruiting Email
    00:35:01 LinkedIn Adds AI Slop Reporting
    00:36:34 Did LinkedIn Create Its Own AI Slop Problem?
    00:39:47 What A Real LinkedIn Strategy Still Requires
    00:40:55 AI Slop Versus Empty Engagement
    00:43:38 T3 Code And Mobile AI Development
    00:44:34 Jack Dorsey’s Buzz And Multi-Agent Collaboration
    00:46:08 AI Agents Working Together On Shared Projects
    00:47:38 Gemini Robotics And One Brain For Any Robot
    00:48:35 Robots Collaborating With Each Other
    00:50:18 Gemini Security Tools Fix 1,072 Chrome Bugs
    00:51:32 Google’s AI Strategy Beyond Frontier Chatbots
    00:53:00 Gemini 3.1 Pro, 3.5 And What Comes Next
    00:55:47 AI Security Models And Finding New Bugs
    00:57:27 Website, Community And Merch Discussion
    00:58:57 Episode Wrap-Up And Three-Year Anniversary

    The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons.
More Technology podcasts
About The Daily AI Show
The Daily AI Show is a panel discussion hosted LIVE each weekday at 10am Eastern. We cover all the AI topics and use cases that are important to today's busy professional. No fluff. Just 45+ minutes to cover the AI news, stories, and knowledge you need to know as a business professional. About the crew: We are a group of professionals who work in various industries and have either deployed AI in our own environments or are actively coaching, consulting, and teaching AI best practices. Your hosts are: Brian Maucere Beth Lyons Andy Halliday Jyunmi Hatcher Karl Yeh
Podcast website

Listen to The Daily AI Show, Dwarkesh Podcast and many other podcasts from around the world with the radio.net app

Get the free radio.net app

  • Stations and podcasts to bookmark
  • Stream via Wi-Fi or Bluetooth
  • Supports Carplay & Android Auto
  • Many other app features