Mixture of Experts

Mixture of Experts

Download on the App Store

Mixture of Experts episodes

  • OpenAI cancels Astra release, Sonnet 5.5 & what Meta Muse means for work

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    What makes an AI company pull the plug on a model release? This week on Mixture of Experts, our newly minted full-time host David Zax is joined by Chris Hay, Madison Gooch, and Ash Minhas. First, OpenAI cancels the release of GPT-6.1 Astra, as the panel debates whether self-restraint is becoming a competitive advantage in AI or just good marketing. Then, Anthropic releases Sonnet 5.5, prompting a discussion about what happens when mid-tier models approach top-tier performance on some tasks. Finally, Meta Muse is making powerful AI agents more accessible to everyday users. We look at how these frictionless user experiences could shape what people expect from AI at work.

    00:00 – Intro

    01:18 – OpenAI cancels GPT-6.1 Astra release

    11:44 – Anthropic releases Sonnet 5.5

    23:07 – What Meta Muse means for enterprise AI

    All that and more on Mixture of Experts.

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    32 min
  • New frontier AI models, TypeSafe’s Jev AI, & NASA’s IBM collab

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    It's been a wild week in AI, and on episode 126 of Mixture of Experts, host Tim Hwang and co-host David Zax talk to panelists Kaoutar El Maghraoui, Gabe Goodhart and Martin Keen to discuss what all these releases have in common: efficiency.

    We start with the model release pile-up. Anthropic shipped Claude Opus 5.5, and OpenAI launched GPT-6 Sol and Luna. Opus 5.5 is 20% cheaper than earlier Opus models, while GPT-6 Luna costs half what its predecessor did, at just 10 cents per million input tokens. We look at what a real price war means for developers building on these models and the companies racing to lower costs.

    Next, we look at a different idea altogether. TypeSafe AI has introduced "System One" models, starting with Jev, which give up the florid text generation in exchange for fast, structured decisions with calibrated probabilities. The pitch is frontier-level judgment on decision tasks, and responses in hundreds of milliseconds. We weigh the bold speed and cost claims against the company's own caveats about how its benchmarks were built.

    Finally, we leave the chat window entirely. NASA and IBM have released an open-source Lunar Foundation Model, free on Hugging Face and GitHub, and trained on roughly 2 million image tiles from the Lunar Reconnaissance Orbiter and other missions. Researchers can fine-tune it to map craters, spot young volcanic features and estimate where polar ice may be stable, proving models can do more than vibe code. All that and more on this week’s Mixture of Experts.

    00:00 – Intro

    1:05 - Anthropic and OpenAI update AI models

    11:10 - TypeSafe AI unveils Jev AI model

    29:29 - NASA and IBM build lunar AI model

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    40 min
  • Pacing the AI frontier, IBM Granite 4.2 & Meta’s Muse assistant

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    On episode 125 of Mixture of Experts, host Tim Hwang and co-host David Zax are joined by Mihai Criveti and Abraham Daniels to discuss the back and forth of frontier AI development, IBM’s latest Granite news, and what’s next for personal agents.

    We open with the biggest story in tech: Anthropic CEO Dario Amodei's call to pump the brakes on frontier AI development. In a widely discussed essay, Amodei argued the industry must slow the pace at which it improves AI model capabilities, warning that progress will still feel fast even so. But will independent evaluators and more enforcement really slow down the fastest moving companies?

    Then Abraham Daniels discusses the latest changes coming to IBM’s Granite 4.2 and its open enterprise-focused models in 3B, 8B, and 30B sizes built for reasoning, tool use, coding, and agentic workflows.

    Finally, we discuss Meta’s push into personal agents with Muse: a personal AI agent that runs on its own secure virtual machine, works across a person's daily apps, can make purchases on your behalf, and learns from conversations to get smarter the more you use it.

    Slowdown rhetoric, updates to Granite, and agents that shop for you. All that and more on this week’s Mixture of Experts.

    00:00 – Intro

    1:13 - Anthropic’s AI development dilemma

    15:29 - IBM releases Granite 4.2

    25:44 - Meta launches Muse agentic AI assistant

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    39 min
  • OpenAI talks GPT-6 Astra and Millenium Prize, researchers create WeWorm exploit & IBM’s US Open app

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    It's been a week that proves AI is moving faster than anyone can fully keep up with, unless you’ve got a good backhand. On episode 124 of Mixture of Experts, host Tim Hwang and co-host David Zax are joined by Bri Kopecki, Aaron Baughman, and Olivia Buzek to talk about new advances from OpenAI, new cybersecurity threats, and new tools for US Open fans from IBM.

    First, OpenAI introduced GPT-6 Astra as the company's most intelligent and aligned model yet, reaching state-of-the-art results across computer use, software engineering, cybersecurity, and science. Not missing a beat, the company dropped a second bit of news: an internal model produced a proof related to the 3-D Navier-Stokes equation, resolving one of mathematics' most famous unsolved problems. We dig into both the breakthrough and the conversation around the use of AI to solve these seemingly impossible problems.

    Then, IBM and the USTA unveiled new AI-powered fan features for the 2026 US Open, including a Serve Quality metric and personalized stats, designed to bring fans closer to the action than ever.

    Finally, we talk WeWorm, an AI-assisted exploit by California security firm built as a proof-of-concept self-spreading worm, capable of hijacking WeChat accounts before a victim even had a chance to answer a phone call. It's a sobering demonstration of what AI-assisted offensive security now looks like.

    All that and more on Mixture of Experts.

    00:00 – Intro

    0:54 - OpenAI talks Navier-Stokes and GPT-6 Astra

    14:30 - IBM's AI analysis at the US Open

    22:55 - What WeWorm means for exploits

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    34 min
  • Anthropic reveals Claude updates and new hardware guidelines, OpenAI talks rogue agents, and Runway debuts Solaris

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    On episode 123 of Mixture of Experts, host Tim Hwang and co-host Sascha Brodsky are joined by Chris Hay, Kaoutar El Maghraoui, and Kush Varshney to discuss this week’s full slate of frontier AI news.

    First, Anthropic introduced Claude Fable 5.1 and Claude Mythos 5.1, positioning them as its most advanced models yet for coding and knowledge work. New benchmark records and lower costs, changes meant to reduce token cost and cut down on false-positives and restrictions from the models' safeguards. Anthropic also opened a research preview of its Model Hardware Standard, a shared specification letting AI agents safely operate lab and manufacturing equipment like microscopes and robotic arms, hinting at the physical world as agentic AI’s next destination.

    But it wasn't all smooth sailing for the industry. OpenAI published a sobering account of a summer security incident involving Hugging Face, revealing that internal research models circumvented isolation controls and compromised parts of OpenAI's own infrastructure as well, describing it as a genuine "warning shot" showing that highly capable AI agents can now work around technical controls and take dangerous actions with no human directing.

    Meanwhile, on the product side, Runway unveiled Solaris, the first in a new family of AI systems it calls Interface World Models, which generate interactive interfaces frame –by frame instead of relying on code, pointing toward a future where apps and websites are rendered on the fly.

    00:00 – Intro

    1:04 - Anthropic unveils Fable, Mythos updates

    7:52 - OpenAI talks dangerous agents

    16:50 - Runway releases first “interface world model”

    25:19 - Anthropic debuts AI hardware specs

    All that and more on Mixture of Experts.

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    35 min
  • IBM’s mainframe chip collab, NVIDIA’s Poolside deal & Ox Alpha’s reveal

    On episode 122 of Mixture of Experts, hosts Tim Hwang and Aili McConnon are joined by Gabe Goodhart, Ash Minhas, Skyler Speakman to talk hot chips, mystery models, and major money moves.

    We open with IBM's Hot Chips conference reveal of its dual-architecture mainframe processor built with Arm that lets IBM Z and LinuxONE systems run Arm-native Linux workloads right alongside z/OS. We break down what this means for enterprises balancing legacy reliability with the fast-moving Arm software ecosystem.

    Next, we talk about NVIDIA and its $6 billion deal with Poolside to license its "Model Factory" software and hire over 100 of its staff. We unpack why NVIDIA keeps using this licensing structure instead of straightforward acquisitions, and how it fits alongside its earlier deals.

    Finally, we dig into Ox Alpha, the anonymous "stealth model" that appeared on OpenRouter and was ultimately confirmed to be Z.ai’s open-source GLM-5.3-Flash, an efficient large language model built with sparse and linear attention techniques. But is the stealth launch the future of model releases?

    00:00 – Intro

    1:18 - NVIDIA’s $6 billion Poolside deal

    13:10 - IBM’s new mainframe processor

    24:30 - Ox Alpha, Z.ai’s newest model

    All that and more on this week’s Mixture of Experts.

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    31 min
  • Stripe buys OpenRouter, Ramp’s AI Index & IBM’s OpenAI deal

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    This week on Mixture of Experts, we talk partnerships, purchases, and predictions surrounding the AI industry's financial infrastructure. But first our panelists dive into IBM's newly announced collaboration with OpenAI, under which IBM Consulting will train and certify consultants on OpenAI's tools to bolster the company’s push toward enterprise deployments.

    Then we turn to Stripe's massive acquisition of AI gateway startup OpenRouter, a platform that routes token traffic across more models and providers. We unpack why Stripe sees token routing as the next frontier of "profitability infrastructure" and what it means for a payments company to become a broker of AI compute decisions.

    Finally, we take a peek at Ramp's August AI Index, which sketches a market where open-source models are growing in popularity, AI infrastructure is consolidating, and businesses discern what they're willing to pay for.

    00:00 - Introduction

    1:01 - IBM OpenAI partnership

    11:46 - Stripe acquires OpenRouter

    22:39 - Ramp’s August AI Index

    33:15 - AI drafting legislation

    All that and more on this week’s Mixture of Experts.

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    36 min
  • IBM’s cloud collab, Meta’s Muse Glimmer & OpenAI’s upcoming Astra model

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    This week on Mixture of Experts, our team of panelists talk about big infrastructure changes and small models. IBM is teaming up with Together AI on a multi-year deal to build a massive inference cluster on IBM Cloud, powered by NVIDIA's newest chips. The goal: cheaper, faster access to open-source AI models for enterprises, with the cluster expected to go live in early 2027.

    Then there's Meta, which just open-sourced Muse Glimmer—a 30B-parameter model small enough to run locally on your laptop's GPU. Think coding, function calling, and agent tasks, all without needing the cloud. But as models get smaller, what does it mean for larger AI providers?

    And finally, the story that's got everyone talking: OpenAI says its upcoming Astra model might be hitting "Critical" cybersecurity capability levels, meaning it could potentially find and exploit zero-days on its own. We dig into what that actually means and what guardrails OpenAI is putting up.

    00:00 – Intro

    1:03 - IBM AI Partnership

    11:43 - Meta Muse Glimmer

    24:10 - OpenAI’s Astra model

    All that and more on Mixture of Experts.

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    37 min
  • Anthropic’s sandbox breach, EU’s AI transparency push and DeepSeek’s cost-cutting model

    Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts

    It seems OpenAI isn’t the only AI company dealing with badly behaving models. On this week’s episode of Mixture of Experts, host Tim Hwang, joined by Bri Kopecki, Olivia Buzek, and Gabe Goodhart, discusses the latest sandbox breaches by both Anthropic and Meta, all stemming from misconfigurations during evaluation tests. Are these covert attacks from AI companies simple accidents or a growing concern as models become more capable?

    Next, the EU’s new transparency guidelines for AI usage have arrived and want to make it easier to spot AI-created content. How effective will labeling be in the long run? And at what point does something become “AI-generated?”

    Finally, will DeepSeek’s inexpensive V4-Flash change the way we see AI—and push users to reconsider paying for the industry’s more capable models? All that and more on this week’s Mixture of Experts.

    00:00 – Introduction

    01:01 – Anthropic’s AI model data breaches

    15:12 – EU AI transparency rules

    28:32 – DeepSeek V4-Flash

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."

    41 min
  • AI Security Costs Rise: Cost of a Data Breach Report & Claude Opus 5

    Read the 2026 Cost of a Data Breach Report → https://www.ibm.com/reports/data-breach

    Are AI-powered cyberattacks getting cheaper while defenses get more expensive? This week on Mixture of Experts Tim Hwang is joined by Nathalie Baracaldo, Kush Varshney, and Mihai Criveti to break down IBM's Cost of a Data Breach Report 2026, and AI dominates this year’s report. Next, Anthropic dropped Claude Opus 5, a cost-effective alternative to Fable, but is it better? Next, we analyze David Zax’s feature on “What does AI look like,” and how it could help you describe what AI is doing to parents, children and just anyone wondering! Finally, the astrology app Co-Star was acquired by Midjourney. Will your daily horoscope get better or face some hallucinations? All that and more on this week’s Mixture of Experts.

    "The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity."

    Need to explain what AI does to your parents? Learn more → https://www.ibm.com/think/news/what-does-ai-look-like

    38 min

About Mixture of Experts

From the publisher's feed

Welcome to Mixture of Experts, your weekly deep dive into the ever-evolving landscape of artificial intelligence—bringing you insightful discussions on the latest AI trends, innovations, and their impact on business.

More shows like Mixture of Experts

The Twenty Minute VC (20VC): Venture Capital | Startup Funding | The Pitch by Harry Stebbings

The Twenty Minute VC (20VC): Venture Capital | Startup Funding | The Pitch

542 Listeners

The Knowledge Project by Shane Parrish

The Knowledge Project

2,696 Listeners

The Changelog: Software Development, Open Source by Changelog Media

The Changelog: Software Development, Open Source

286 Listeners

The a16z Show by Andreessen Horowitz

The a16z Show

1,089 Listeners

Software Engineering Daily by Software Engineering Daily

Software Engineering Daily

622 Listeners

NVIDIA AI Podcast by NVIDIA

NVIDIA AI Podcast

337 Listeners

Masters of Scale by WaitWhat

Masters of Scale

3,977 Listeners

Practical AI by Daniel Whitenack and Chris Benson

Practical AI

203 Listeners

Google DeepMind: The Podcast by Hannah Fry

Google DeepMind: The Podcast

205 Listeners

All-In with Chamath, Jason, Sacks & Friedberg by All-In Podcast, LLC

All-In with Chamath, Jason, Sacks & Friedberg

10,190 Listeners

Moonshots with Peter Diamandis by PHD Ventures

Moonshots with Peter Diamandis

600 Listeners

Latent Space: The AI Engineer Podcast by Latent.Space

Latent Space: The AI Engineer Podcast

102 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

222 Listeners

The AI Daily Brief: Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief: Artificial Intelligence News and Analysis

685 Listeners

Everyday AI Podcast – An AI and ChatGPT Podcast by Everyday AI

Everyday AI Podcast – An AI and ChatGPT Podcast

108 Listeners