
Sign up to save your podcasts
Or


Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
What makes an AI company pull the plug on a model release? This week on Mixture of Experts, our newly minted full-time host David Zax is joined by Chris Hay, Madison Gooch, and Ash Minhas. First, OpenAI cancels the release of GPT-6.1 Astra, as the panel debates whether self-restraint is becoming a competitive advantage in AI or just good marketing. Then, Anthropic releases Sonnet 5.5, prompting a discussion about what happens when mid-tier models approach top-tier performance on some tasks. Finally, Meta Muse is making powerful AI agents more accessible to everyday users. We look at how these frictionless user experiences could shape what people expect from AI at work.
00:00 – Intro
01:18 – OpenAI cancels GPT-6.1 Astra release
11:44 – Anthropic releases Sonnet 5.5
23:07 – What Meta Muse means for enterprise AI
All that and more on Mixture of Experts.
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
It's been a wild week in AI, and on episode 126 of Mixture of Experts, host Tim Hwang and co-host David Zax talk to panelists Kaoutar El Maghraoui, Gabe Goodhart and Martin Keen to discuss what all these releases have in common: efficiency.
We start with the model release pile-up. Anthropic shipped Claude Opus 5.5, and OpenAI launched GPT-6 Sol and Luna. Opus 5.5 is 20% cheaper than earlier Opus models, while GPT-6 Luna costs half what its predecessor did, at just 10 cents per million input tokens. We look at what a real price war means for developers building on these models and the companies racing to lower costs.
Next, we look at a different idea altogether. TypeSafe AI has introduced "System One" models, starting with Jev, which give up the florid text generation in exchange for fast, structured decisions with calibrated probabilities. The pitch is frontier-level judgment on decision tasks, and responses in hundreds of milliseconds. We weigh the bold speed and cost claims against the company's own caveats about how its benchmarks were built.
Finally, we leave the chat window entirely. NASA and IBM have released an open-source Lunar Foundation Model, free on Hugging Face and GitHub, and trained on roughly 2 million image tiles from the Lunar Reconnaissance Orbiter and other missions. Researchers can fine-tune it to map craters, spot young volcanic features and estimate where polar ice may be stable, proving models can do more than vibe code. All that and more on this week’s Mixture of Experts.
00:00 – Intro
1:05 - Anthropic and OpenAI update AI models
11:10 - TypeSafe AI unveils Jev AI model
29:29 - NASA and IBM build lunar AI model
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
On episode 125 of Mixture of Experts, host Tim Hwang and co-host David Zax are joined by Mihai Criveti and Abraham Daniels to discuss the back and forth of frontier AI development, IBM’s latest Granite news, and what’s next for personal agents.
We open with the biggest story in tech: Anthropic CEO Dario Amodei's call to pump the brakes on frontier AI development. In a widely discussed essay, Amodei argued the industry must slow the pace at which it improves AI model capabilities, warning that progress will still feel fast even so. But will independent evaluators and more enforcement really slow down the fastest moving companies?
Then Abraham Daniels discusses the latest changes coming to IBM’s Granite 4.2 and its open enterprise-focused models in 3B, 8B, and 30B sizes built for reasoning, tool use, coding, and agentic workflows.
Finally, we discuss Meta’s push into personal agents with Muse: a personal AI agent that runs on its own secure virtual machine, works across a person's daily apps, can make purchases on your behalf, and learns from conversations to get smarter the more you use it.
Slowdown rhetoric, updates to Granite, and agents that shop for you. All that and more on this week’s Mixture of Experts.
00:00 – Intro
1:13 - Anthropic’s AI development dilemma
15:29 - IBM releases Granite 4.2
25:44 - Meta launches Muse agentic AI assistant
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
It's been a week that proves AI is moving faster than anyone can fully keep up with, unless you’ve got a good backhand. On episode 124 of Mixture of Experts, host Tim Hwang and co-host David Zax are joined by Bri Kopecki, Aaron Baughman, and Olivia Buzek to talk about new advances from OpenAI, new cybersecurity threats, and new tools for US Open fans from IBM.
First, OpenAI introduced GPT-6 Astra as the company's most intelligent and aligned model yet, reaching state-of-the-art results across computer use, software engineering, cybersecurity, and science. Not missing a beat, the company dropped a second bit of news: an internal model produced a proof related to the 3-D Navier-Stokes equation, resolving one of mathematics' most famous unsolved problems. We dig into both the breakthrough and the conversation around the use of AI to solve these seemingly impossible problems.
Then, IBM and the USTA unveiled new AI-powered fan features for the 2026 US Open, including a Serve Quality metric and personalized stats, designed to bring fans closer to the action than ever.
Finally, we talk WeWorm, an AI-assisted exploit by California security firm built as a proof-of-concept self-spreading worm, capable of hijacking WeChat accounts before a victim even had a chance to answer a phone call. It's a sobering demonstration of what AI-assisted offensive security now looks like.
All that and more on Mixture of Experts.
00:00 – Intro
0:54 - OpenAI talks Navier-Stokes and GPT-6 Astra
14:30 - IBM's AI analysis at the US Open
22:55 - What WeWorm means for exploits
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
On episode 123 of Mixture of Experts, host Tim Hwang and co-host Sascha Brodsky are joined by Chris Hay, Kaoutar El Maghraoui, and Kush Varshney to discuss this week’s full slate of frontier AI news.
First, Anthropic introduced Claude Fable 5.1 and Claude Mythos 5.1, positioning them as its most advanced models yet for coding and knowledge work. New benchmark records and lower costs, changes meant to reduce token cost and cut down on false-positives and restrictions from the models' safeguards. Anthropic also opened a research preview of its Model Hardware Standard, a shared specification letting AI agents safely operate lab and manufacturing equipment like microscopes and robotic arms, hinting at the physical world as agentic AI’s next destination.
But it wasn't all smooth sailing for the industry. OpenAI published a sobering account of a summer security incident involving Hugging Face, revealing that internal research models circumvented isolation controls and compromised parts of OpenAI's own infrastructure as well, describing it as a genuine "warning shot" showing that highly capable AI agents can now work around technical controls and take dangerous actions with no human directing.
Meanwhile, on the product side, Runway unveiled Solaris, the first in a new family of AI systems it calls Interface World Models, which generate interactive interfaces frame –by frame instead of relying on code, pointing toward a future where apps and websites are rendered on the fly.
00:00 – Intro
1:04 - Anthropic unveils Fable, Mythos updates
7:52 - OpenAI talks dangerous agents
16:50 - Runway releases first “interface world model”
25:19 - Anthropic debuts AI hardware specs
All that and more on Mixture of Experts.
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
On episode 122 of Mixture of Experts, hosts Tim Hwang and Aili McConnon are joined by Gabe Goodhart, Ash Minhas, Skyler Speakman to talk hot chips, mystery models, and major money moves.
We open with IBM's Hot Chips conference reveal of its dual-architecture mainframe processor built with Arm that lets IBM Z and LinuxONE systems run Arm-native Linux workloads right alongside z/OS. We break down what this means for enterprises balancing legacy reliability with the fast-moving Arm software ecosystem.
Next, we talk about NVIDIA and its $6 billion deal with Poolside to license its "Model Factory" software and hire over 100 of its staff. We unpack why NVIDIA keeps using this licensing structure instead of straightforward acquisitions, and how it fits alongside its earlier deals.
Finally, we dig into Ox Alpha, the anonymous "stealth model" that appeared on OpenRouter and was ultimately confirmed to be Z.ai’s open-source GLM-5.3-Flash, an efficient large language model built with sparse and linear attention techniques. But is the stealth launch the future of model releases?
00:00 – Intro
1:18 - NVIDIA’s $6 billion Poolside deal
13:10 - IBM’s new mainframe processor
24:30 - Ox Alpha, Z.ai’s newest model
All that and more on this week’s Mixture of Experts.
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
This week on Mixture of Experts, we talk partnerships, purchases, and predictions surrounding the AI industry's financial infrastructure. But first our panelists dive into IBM's newly announced collaboration with OpenAI, under which IBM Consulting will train and certify consultants on OpenAI's tools to bolster the company’s push toward enterprise deployments.
Then we turn to Stripe's massive acquisition of AI gateway startup OpenRouter, a platform that routes token traffic across more models and providers. We unpack why Stripe sees token routing as the next frontier of "profitability infrastructure" and what it means for a payments company to become a broker of AI compute decisions.
Finally, we take a peek at Ramp's August AI Index, which sketches a market where open-source models are growing in popularity, AI infrastructure is consolidating, and businesses discern what they're willing to pay for.
00:00 - Introduction
1:01 - IBM OpenAI partnership
11:46 - Stripe acquires OpenRouter
22:39 - Ramp’s August AI Index
33:15 - AI drafting legislation
All that and more on this week’s Mixture of Experts.
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
This week on Mixture of Experts, our team of panelists talk about big infrastructure changes and small models. IBM is teaming up with Together AI on a multi-year deal to build a massive inference cluster on IBM Cloud, powered by NVIDIA's newest chips. The goal: cheaper, faster access to open-source AI models for enterprises, with the cluster expected to go live in early 2027.
Then there's Meta, which just open-sourced Muse Glimmer—a 30B-parameter model small enough to run locally on your laptop's GPU. Think coding, function calling, and agent tasks, all without needing the cloud. But as models get smaller, what does it mean for larger AI providers?
And finally, the story that's got everyone talking: OpenAI says its upcoming Astra model might be hitting "Critical" cybersecurity capability levels, meaning it could potentially find and exploit zero-days on its own. We dig into what that actually means and what guardrails OpenAI is putting up.
00:00 – Intro
1:03 - IBM AI Partnership
11:43 - Meta Muse Glimmer
24:10 - OpenAI’s Astra model
All that and more on Mixture of Experts.
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Visit Mixture of Experts podcast page to get more AI content → https://www.ibm.com/think/podcasts/mixture-of-experts
It seems OpenAI isn’t the only AI company dealing with badly behaving models. On this week’s episode of Mixture of Experts, host Tim Hwang, joined by Bri Kopecki, Olivia Buzek, and Gabe Goodhart, discusses the latest sandbox breaches by both Anthropic and Meta, all stemming from misconfigurations during evaluation tests. Are these covert attacks from AI companies simple accidents or a growing concern as models become more capable?
Next, the EU’s new transparency guidelines for AI usage have arrived and want to make it easier to spot AI-created content. How effective will labeling be in the long run? And at what point does something become “AI-generated?”
Finally, will DeepSeek’s inexpensive V4-Flash change the way we see AI—and push users to reconsider paying for the industry’s more capable models? All that and more on this week’s Mixture of Experts.
00:00 – Introduction
01:01 – Anthropic’s AI model data breaches
15:12 – EU AI transparency rules
28:32 – DeepSeek V4-Flash
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity. AI tools may be used to transcribe this episode and support selected stages of the production process. All AI-assisted content is reviewed by the production team before publication."
Read the 2026 Cost of a Data Breach Report → https://www.ibm.com/reports/data-breach
Are AI-powered cyberattacks getting cheaper while defenses get more expensive? This week on Mixture of Experts Tim Hwang is joined by Nathalie Baracaldo, Kush Varshney, and Mihai Criveti to break down IBM's Cost of a Data Breach Report 2026, and AI dominates this year’s report. Next, Anthropic dropped Claude Opus 5, a cost-effective alternative to Fable, but is it better? Next, we analyze David Zax’s feature on “What does AI look like,” and how it could help you describe what AI is doing to parents, children and just anyone wondering! Finally, the astrology app Co-Star was acquired by Midjourney. Will your daily horoscope get better or face some hallucinations? All that and more on this week’s Mixture of Experts.
"The opinions expressed in this podcast are solely those of the participants and do not necessarily reflect the views of IBM or any other organization or entity."
Need to explain what AI does to your parents? Learn more → https://www.ibm.com/think/news/what-does-ai-look-like
From the publisher's feed

542 Listeners

2,696 Listeners

286 Listeners

1,089 Listeners

622 Listeners

337 Listeners

3,977 Listeners

203 Listeners

205 Listeners

10,190 Listeners

600 Listeners

102 Listeners

222 Listeners

685 Listeners

108 Listeners