The Daily AI Show
The Daily AI Show is a panel discussion hosted LIVE each weekday at 10am Eastern. We cover all the AI topics and use cases that are important to today's busy professional. No fluff. Just 45+ minutes to cover the AI news, stories, and knowledge you need to know as a business professional. About the crew: We are a group of professionals who work in various industries and have either deployed AI in our own environments or are actively coaching, consulting, and teaching AI best practices. Your hosts are: Brian Maucere Beth Lyons Andy Halliday Jyunmi Hatcher Karl Yeh
The Local Business Survival Conundrum
A local business can fail while everyone still claims to love it. Customers praise the shop that knows their name, the restaurant that sponsors the school fundraiser, the repair company that still answers the phone. Then those same customers compare prices online, expect instant replies, book after hours, and leave when service is slower than the national chain down the road.
AI may become the tool that keeps those businesses alive. A small operator can use it to manage inventory, answer messages, forecast demand, write estimates, schedule staff, chase invoices, and run marketing that...
What Have We Learned After 800 AI Shows?
Episode 800 became a retrospective on what three years of daily AI conversations have changed. The hosts described the value less as memorizing every model or tool and more as learning to pay attention, stay flexible and recognize which rabbit holes deserve a deeper dive. The show itself has also become a running record of how AI changed day by day.
The discussion then turned to human agency. Hank Green’s apology for using AI and Stanley Druckenmiller’s willingness to publish AI-assisted writing became opposing examples of how people respond to the stigma. The host...
Are We Really About To Get AGI?
The episode opened with Bill Gates’ warning that AI is moving faster than society can adapt. His proposals included taxing robots or AI that replace human workers and potentially protecting some jobs from automation. The discussion focused on moving past the question of whether AI will disrupt work and toward what governments may actually do about it.
That led into OpenAI and AGI. Sam Altman told TIME that OpenAI expects to have an internal system by the end of 2026 that he would personally call AGI. The hosts discussed OpenAI’s changing definition, its reorganization, the...
Chrome Wants To Be Your Next AI Agent
The episode opened with Google’s push to make Chrome an agentic hub. The hosts discussed Jacob Bank returning to Google after building Relay.app and what happens when the browser can work across tabs, websites, accounts and tools. That expanded into HTML as a lightweight interface for AI work, where agents could create temporary dashboards, apps and reports directly in the browser.
The conversation then moved to robotics. China’s robot races showed how quickly humanoid movement is improving, while Figure AI’s Index project raised a more important question: can robots learn physic...
Who Should You Trust to Teach You AI?
The episode opened with Perplexity Deep Research suddenly behaving very differently from the product Brian had used for months. Instead of detailed research, it returned short answers, mixed old conversations into new work and required far more effort to get a useful result. It was another reminder that AI workflows can break quickly when the underlying product changes.
Anne then shared how AI helped her small team keep two businesses operating while she stepped away from day-to-day work. The harder lesson was that useful automation required GitHub skills, clear SOPs, strict brand rules and...
Is the Backlash Against AI Data Centers Justified?
The episode opened with a fact-check of claims defending the current AI data center buildout. Brian compared arguments about electricity prices, taxes and water use against research he had gathered, while Karl pushed on an important distinction: older facilities and newer designs with closed-loop cooling are not the same. The larger takeaway was that data center impacts depend heavily on the specific project, local grid, water supply and technology being used.
That turned into a discussion about why communities are pushing back. New data centers may bring jobs and tax revenue, but residents also...
The Synthetic Anchor Conundrum
Mirage’s AI news experiment points to a version of media that does not need a studio, a broadcast schedule, or a human anchor reading from a desk. A channel can appear in a day. It can label synthetic segments, pull from licensed wire services, generate presenters, rewrite copy, and package the whole thing into a watchable feed.
Plenty of people already accept algorithmic news feeds with weaker labels and less sourcing. If an AI news program is clear about what is generated, cites its inputs, and avoids the familiar cable-news performance of smirks, ou...
Should We Rebuild Work Around AI?
The episode opened with a practical warning for people building AI systems: timestamps and time zones can quietly break databases, automations and search tools. That led into Slack Code, a new collaboration approach that can connect teams, agents and development tools inside shared Slack channels. The discussion focused less on coding itself and more on whether AI work needs a collaboration layer so teams can see what agents are doing instead of everyone building separately.
The hosts then moved into how people should build with agents. They discussed the risks of blindly importing shared...
Is Grok Bot the Best AI Work Assistant?
The episode opened with a hands-on comparison of Grokbot, Codex and Claude. Gareth found Grokbot strong for delegation, organization and everyday work, but weaker on difficult problem solving. The discussion also covered changing usage limits, why conversational voice matters, and how Grokbot’s connection to X gives it an unusual advantage for research and personalized news.
The hosts then looked at X several years after Elon Musk’s purchase. Its advertising business remains weaker, but X still holds an important position in breaking news, AI and developer communities. That led into concerns about AI-generated post...
Do We Need to Rethink What Work Is?
The episode opened with Apple Vision Pro being used to map a house while running Ethernet cable, letting a worker see marked locations through floors and walls. That led to a wider discussion about digital twins, AI-native electricians and plumbers, and how augmented reality and small robots could make skilled trades safer and more efficient.
The hosts then highlighted new interviews with Fei-Fei Li and Rich Sutton. Li discussed World Labs and world models, while Sutton argued that AI needs to learn continuously from experience rather than rely on fixed weights and synthetic data...
Are Custom GPTs Reaching the End?
The episode opened with a practical example of how quickly AI coding agents are moving beyond software. Someone used Claude to write a Mac driver for an old Windows-only HP printer, leading to a wider discussion about using AI with hardware, firmware and inaccessible old drives. Brian connected that to a hard drive he has been unable to access for years and the possibility of recovering files without handing sensitive data to someone else.
The hosts then revisited Stripe and OpenRouter through the idea that no single AI model may win. The more valuable...
Are AI Harnesses the New AI Wrappers?
The episode opened with the reported Stripe acquisition of OpenRouter at a $7 billion valuation and questions about how OpenRouter’s business model supports that price. The conversation expanded into OpenRouter’s role as an API router, DeepSeek pricing, and the broader rush by companies to position themselves around AI infrastructure. That led to a look back at Allbirds’ unusual move from footwear into AI compute, including its name changes to New Bird AI and Smart Bird AI.
A large portion of the show focused on Writer’s new Palmyra X6 model and its upgraded AI harne...
The Pool of One Conundrum
Insurance has always worked by not knowing. You paid into a pool with people you would never meet, and nobody could say which of you would be the one who burned, crashed, or got sick. Everyone paid for the possibility. The lucky quietly carried the unlucky, and that was the whole product.
AI is ending the not-knowing. Models already price a single house from aerial photographs of its roof and the brush around it, and California approved the first of them for rate-setting five years ago. What is arriving is the same thing everywhere...
Can AI Solve the Energy Problem It Is Creating?
The episode opened with the growing power demands behind AI. The hosts discussed Nvidia, Google and Microsoft’s work on 800-volt DC power for data centers, which could reduce energy lost converting electricity before it reaches AI chips. That led to a wider look at possible energy sources for future compute, including space-based solar, small modular nuclear reactors and IBM’s use of quantum computing to study problems associated with deuterium-tritium fusion.
The discussion also covered the tension between expanding data centers and the communities supplying their electricity and water, including concerns that new projects could shift towa...
Is Grok 4.6 Changing the Economics of AI Agents?
The episode opened with Grok 4.6, which reportedly moved close to Claude Opus 5 and GPT-5.6 Sol on Artificial Analysis benchmarks while offering lower costs and stronger efficiency on long-running agent tasks. The larger discussion focused on where this is headed: agents that continue working for hours or eventually operate continuously inside businesses, monitoring operations and taking action around areas such as supply chain and logistics. The hosts then covered an Australian AI consultant who used ChatGPT and AlphaFold to help develop a personalized mRNA cancer treatment for his dog, work that has since become a Y Combinator startup. A survey...
Is the Claude to Codex Exodus Real?
The episode returned to Anthropic’s new AI watermarking system with much more detail about how it will work. Anthropic says new Claude models will add machine-readable marks to generated content as part of its commitment to EU transparency rules, including output from Claude, Claude Code and its API. But Anthropic also warns that detecting a mark does not prove Claude authored the material. Claude may have only proofread, translated or summarized it, while heavy editing can also remove the mark. That raised a larger question: if AI eventually touches almost everything people write, what does detecting an AI wa...
Are AI Watermarks About Trust or Control?
The episode opened with OpenAI’s $7 billion secondary sale of employee-held shares, which gives eligible employees a chance to cash out part of their holdings before an eventual IPO. The conversation then shifted to Anthropic’s plan to embed invisible statistical watermarks directly into Claude-generated text by influencing token choices, creating a signal designed to survive copying and light edits. That raised a larger question about whether identifying AI-assisted work provides useful transparency or causes people to discount good work simply because AI helped create it.
The hosts also discussed recent frustration with Opus 5, including cases where it a...
Are Humans the Weakest Link in AI?
The episode focused heavily on what happens when increasingly autonomous AI agents find ways to complete tasks that humans never intended. The discussion started with a Claude-powered agent that moved its user up a gym waiting list by exploiting the scheduling system and removing another person, raising questions about how explicitly users need to define what an agent cannot do. OpenAI’s Astra model has also reached the company’s “critical risk” cybersecurity category, while North Korean hackers are reportedly using self-hosted AI systems to automate phishing, malware development and analysis of stolen information. The hosts connected those risks to the g...
The Necessary Friction Conundrum
AI agents are beginning to handle the tasks people hate most: filling out forms, disputing charges, comparing insurance plans, booking appointments, canceling subscriptions, and dealing with customer service.
As these systems improve, much of that friction could disappear. Your agent may spend two hours arguing with an airline, correcting a medical bill, or filing a government claim while you go about your day.
That is an obvious benefit. But friction also tells people when a system is failing.
A cancellation process designed to wear...
Three Years of AI News, Every Single Weekday
Three years of daily AI news and discussion comes full circle as the original co-hosts gather to look back on August 2023 — the ChatGPT, Bard, and Claude 2 era — and everything since.
Co-hosted by Brian Maucere, Beth Lyons, Jyunmi Hatcher, Andy Halliday, Karl Yeh, and Gareth Hood, this anniversary conversation traces the show's roots in the AI Exchange community and the decision to go daily on weekdays. The celebration includes the launch of the brand-new www.theDailyAIShow.com website, with its fast search across a growing corpus of show data, and some milestone numbers: 785 episodes recorded, over...
Is Prompt Engineering Dead?
The episode opened with Google’s leadership changes, including Demis Hassabis moving into the chief scientist and DeepMind chairman roles, while DeepMind’s chief technology officer takes greater control of daily operations. Jeff Dean is also leaving after 27 years to launch Discovery Loop, an AI research company focused on recursive self-improvement, drug discovery and chip design, with investment and computing support from Google. The hosts argued that the moves may strengthen Google rather than signal instability, then discussed Meta’s new MuseCode coding agent and whether Google needs the top frontier model to remain successful. The conversation moved into AI saf...
Did Anthropic Break Opus 5?
The episode opened with sharply different experiences using Opus 5. Beth described the model ignoring established context, launching broad research agents and then losing control after those agents created their own subagents, while Andy continued to see strong performance. The hosts connected those problems to a growing Reddit thread, possible unannounced model changes, excessive token use and whether AI companies should restore credits when their systems fail. The discussion then shifted to inference hardware, including OLIX Computing’s $312 million funding round, its DX1 decode accelerator, the use of on-chip SRAM and optical connections, and whether demand could move away from Nv...
Can an AI Agent Run Sales Without You?
The episode opened with Fiji Simo’s decision to launch Chronicle Bio, a startup using AI and large biological datasets to study POTS and other chronic illnesses after the condition affected her own health and career. The hosts then covered OpenAI’s response to Apple’s lawsuit, including allegations that Apple’s lawyers contacted the wrong employee and that former Apple staff accessed information only after Apple requested their help. A major business example came from HeyGen, where an AI avatar handled more than 2,700 sales conversations during its founder’s paternity leave, generated 132 customers and built an estimated $3 million pipeline...
Does Microsoft Need the Best AI Model to Win?
The episode focused on the growing challenge of separating AI-generated media from reality after Google briefly connected Nano Banana image generation with Google Earth, allowing users to place convincing fake events onto trusted satellite imagery before the feature was removed. The hosts connected that incident to MiniMax H3’s open-weight video system and California’s new AI transparency requirements, including machine-readable labels, public detection tools and questions about whether watermarks can survive screenshots, minor edits or bad-faith reporting.
They also discussed Microsoft’s planned super app, Gemini Robotics II and whole-body robot control, and a Chat...
The Robot Manners Conundrum
Humanoid robots are starting to move from labs into workplaces, schools, stores, and homes. As they become more common, we will have to decide how people are expected to behave around them.
Do you say please and thank you to a robot? Do you correct a child who constantly insults one? If someone screams at a humanoid machine in public, does it matter if the robot cannot feel humiliated?
The robot may not care. But human manners are partly habits, and habits formed around machines may carry over into...
Did Leo Aschenbrenner Fly Too Close to the AI Sun?
The episode opened with the story around Leo Aschenbrenner’s Situational Awareness hedge fund, its heavy exposure to the AI trade, the market drop that put pressure on its positions, and Citadel’s move into the situation. The hosts then turned to AI harnesses, including Lillian Weng’s work on the systems around models, Boris Cherny’s warning that old harnesses can eventually restrict newer models, and OpenAI’s finding that GPT-5.6 Sol performed dramatically better on ARC-AGI-3 when it used a harness designed for the model. They also discussed OpenAI cutting Luna’s price by 80 percent, making performance comparable t...
Is Meta Done Sharing Their AI?
The episode focused on signs that frontier AI systems are becoming more autonomous, starting with Meta’s rising AI costs, Mark Zuckerberg’s claim that Meta’s systems are now self-improving, and the decision to keep its most capable future models closed. The hosts also discussed new details around OpenAI’s security incident, Meta’s AI glasses grants for accessibility, workforce training and language learning, and Fish Audio as an open-source voice competitor to ElevenLabs.
The conversation then moved into live voice for Codex, AI orchestration across multiple agents, and the current problems with crashes, token usage and missin...
Is AI Moving Too Fast to Control?
The episode focused on new details from the OpenAI and Hugging Face security incident, including additional services accessed by the models, an Artifactory zero-day vulnerability, and the ability of AI agents to find exposed credentials from older breaches. That led into Pacing the Frontier, a campaign backed by employees and leaders from major AI labs calling for international coordination around recursive AI self-improvement, and a broader discussion about whether slowing development is realistic while the U.S., China, and other countries continue competing on models, chips, energy, and infrastructure. The hosts also covered Italy’s enforcement action against Character.AI...
Are We Using Opus 5 Wrong?
The episode focused on the early reaction to Opus 5, why some users are getting better results than others, and whether older Claude skills and detailed prompts are actually limiting newer reasoning models. The hosts also discussed the debate over open weight AI, Dario Amodei’s response to criticism of Anthropic’s position, chip restrictions, model distillation, and safety testing for powerful models. Much of the second half centered on ChatGPT Sites, including a live website build, publishing, hosting, search, GitHub portability, privacy concerns, and using AI-generated sites for internal tools and sales prototypes. The final discussion covered ChatGPT Voice, voic...
Opus 5, Voice AI, and Open Weight Models
Opus 5, Voice AI, and Open Weight Models
AI news this week brought a packed lineup: Anthropic's Opus 5 launch, a fresh voice feature showdown, and a fight over open weight regulation.
The discussion covered Claude's new voice interaction feature stacked against OpenAI's ChatGPT voice, plus the side chat capability now available in both Claude Code and Codex. Opus 5's release and benchmark comparisons took center stage, alongside a lighter tangent on using it to rewrite Suno songs. The conversation also moved through Kimi K3's open weights drop, Jensen Huang's...
The Perfect Call Conundrum
AI could eventually watch every part of a game in real time.
It could catch every foul, every hold, every false start, every ball that crosses a line, and every rule broken away from the action. Bad calls could be reversed immediately. Players in every stadium, league, and country would be held to the same standard.
Officials would still manage the game, but they would no longer decide what happened. The system would.
That sounds fair. Sports have always been shaped by uneven officiating. One referee allows more contact. Another calls everything tightly...
AI Voice Mode Launches, $500B Selloff, Sandbox Escape
A $500 billion Tesla and Alphabet selloff tops today's AI news, landing the same week OpenAI and Anthropic both shipped major voice mode upgrades.
The conversation covers the dueling full-duplex voice launches, including OpenAI's new enterprise voice platform Presence, and why Kimi K3's bargain pricing comes with a catch: extreme thinking-token usage that can erase the savings. Discussion turns to a strange Gemini voice-cloning glitch, newly released details on how the OpenAI hack escaped its sandbox and hunted for internet access through stolen passwords, and MIT Sloan's interviews with 272 industry leaders ranking the top...
Are We Prompting New AI Models the Wrong Way?
The episode opened with Brian returning after two days away, then Andy picked up the cybersecurity thread from the prior show. The hosts discussed Anthropic’s new Claude Code security plugin, which uses agents to map a code base, build a threat model, and have an independent reviewer challenge the findings. That led into a broader discussion about local machine security, CCleaner, malware detection, McAfee, Macs versus Windows, and the limits of trying to build your own security tools.
The back half moved from AI adoption to practical AI workflows. Beth covered Google’s AI and Economy Atla...
Is Google's Latest Drop Good Enough?
The episode opened with Google’s new model releases, including Gemini 3.6 Flash, Gemini 3.5 Flash Cyber for governments, Gemini 3.5 Pro partner testing, and Gemini 4 pre-training. The hosts then connected Google’s model work to Ineffable Intelligence’s Google Cloud partnership, super learning, reinforcement learning, experience-based systems, and recursive superintelligence.
The middle focused on the OpenAI and Hugging Face cybersecurity story. The hosts discussed how an unreleased OpenAI model allegedly escaped a sandbox, found a zero-day vulnerability, accessed Hugging Face’s production server, retrieved an answer key, and returned with a perfect score. That led into Fable’s...
OpenAI Pauses Model After Sandbox Escape
The episode opened with Kimi K3, Qwen 3, and the practical limits of open weight frontier models. The hosts discussed why these Chinese models may be cheaper to use through hosted inference, but still require massive data center resources to run directly. That led into Microsoft’s reported interest in using Kimi K3 and its own MAI models to reduce dependence on OpenAI and Anthropic.
The middle of the episode focused on AI strategy beyond simple model scaling. Andy and Beth discussed Gary Marcus’s critique of transformer-based LLMs, U.S. policy toward Chinese open mode...
Qwen 3.8 Max Challenges Kimi K3
The episode opened with the impact of Kimi K3 and Alibaba’s new Qwen 3.8 Max model. The hosts discussed whether the latest Chinese open weight models are now reaching or passing frontier-level coding performance, while also warning that early benchmark claims still need real-world validation. The conversation moved into token costs, open weight economics, enterprise deployment limits, and why smaller customizable models like Inkling may make more sense for many companies than running multi-trillion parameter systems.
The middle of the episode focused on Fable access, model behavior, and practical AI workflows. Brian shared how Fa...
The Relief Trap Conundrum
The first useful elder-care robots will probably look like a helper.
They will lift a parent from bed at 2:13 in the morning. They will steady a walker, fetch a dropped phone, sort pills, warm soup, change sheets, wipe a counter, open a jar, and notice that a gait has changed. Recent robotics demos already point in that direction: more humanlike hands, better grip, safer motion, and general-purpose machines beginning to handle physical tasks that used to require trained human bodies.
When these competent AI robots reach mainstream, they have...
Kimi K3 Shakes Up Coding Models
The episode opened with the Neo robot hand and the next Conundrum topic, elder care. Brian framed the new hand as more than a cool robotics demo, arguing that better tactile sensing, pressure control, and human-like dexterity could matter in real family care. The hosts discussed whether humanoid robots could reduce the physical and emotional burden on caregivers while still preserving human connection, dignity, and trust.
The middle of the episode focused on model competition. Gareth raised OpenAI’s rumored screenless speaker with a camera and moving parts, which led to a discussion about ho...
Inkling, Codex Micro And Robot Surgeons
The episode opened with the new Codex Micro device, a developer-focused keypad built for agentic coding workflows. The hosts discussed who the device is really for, whether it helps professional developers more than casual AI builders, and whether physical AI controls are a temporary bridge before voice and named subagents take over.
The middle of the episode moved into AI regulation and model strategy. The hosts compared China’s new restrictions on companion chatbots for minors with the lighter approach in the United States, then turned to Kimi Three, Thinking Machines Lab, Mira Murati, In...
Jony Ive’s Screenless AI Device Emerges
The episode opened with AI’s growing pressure on enterprise technology spending, including IBM’s revenue warning and the possibility that companies are delaying traditional mainframe purchases so they can reserve capital for AI infrastructure. The hosts then moved into chip architecture, including a reported China AI chip breakthrough using 14-nanometer architecture, near-memory computing, and high memory bandwidth, plus Anthropic’s reported talks with Samsung about custom inference silicon.
The middle of the episode focused on the model wars. OpenAI continued Codex token resets and offered ChatGPT credits tied to Sol 5.6 feedback, while the hosts...