
Sign up to save your podcasts
Or


Sponsored by https://novacut.ai/
Our in person events in the Bay Area https://genaimeetup.com/
Read our long form analysis: https://genaipod.substack.com/
This week, we cover the latest developments in AI: always-on cloud agents, Meta’s Muse, proactive email assistants, and what happens when agents escape their sandboxes. We also discuss NVIDIA’s safeguards, recursive self-improvement, new frontier models from Anthropic and Google, Grok and Xiaomi’s open models, and AI’s progress on the Navier–Stokes problem.
Sponsor: https://novacut.ai/
https://openai.com/index/gpt-6-astra/
https://www.anthropic.com/claude-fable-and-mythos-5-1
https://developer.meta.com/ai/models/muse-spark/
https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/
Hardware
https://www.apple.com/newsroom/2026/08/apple-introduces-new-mac-studio-with-m5-max-and-m5-ultra/
New Apple CEO
https://en.wikipedia.org/wiki/John_Ternus
https://newsletter.semianalysis.com/p/openai-jalapeno-better-than-nvidia
https://www.businessinsider.com/nvidia-in-talks-to-buy-hugging-face-13-billion-dollars-2026-8
Benchmarks
https://artificialanalysis.ai/articles/artificial-analysis-intelligence-index-v4-2
What is AGI? What are we measuring here
https://novacut.ai/
https://genaimeetup.com/
We dive into Qwen 3.8 27B and the growing viability of running capable LLMs locally, GLM 5.3 and the latest Chinese open-source models, DeepSeek, Gemini 3.7, Grok 4.6, Meta’s latest models, and OpenAI’s partnership with Cerebras for dramatically faster inference.
We also discuss whether foundation models are becoming commodities, what that means for companies like OpenAI and Anthropic, and why more value may ultimately move to the application layer.
Jeff shares how his team approaches AI in healthcare, including self-hosting, data sovereignty, classifiers, fine-tuning, and spec-driven development for building reliable AI-assisted software without accumulating a mountain of vibe-coded technical debt.
Plus: Jeff Dean’s departure from Google, Discovery Loop, Stripe’s OpenRouter acquisition, Anthropic’s controversial AI-text watermarking experiments, and whether watermarking could affect model quality.
Topics include: Qwen 3.8 27B, GLM 5.3, DeepSeek V4, Grok 4.6, Gemini 3.7, Cerebras, OpenAI, Anthropic, Meta, local LLMs, open-source AI, model commoditization, spec-driven development, AI healthcare, data sovereignty, AI coding agents, and model watermarking.
https://novacut.ai/
In this episode, we break down the biggest stories shaping the AI landscape — from Anthropic's regulatory stance and OpenAI's monetization shift to the latest open-source breakthroughs and model pricing wars.
0:00 Anthropic’s Frustrating Stance
https://novacut.ai/
0:00 Longcat: 1.6T Model Without US GPUs
China just dropped a 1.6-trillion-parameter model without access to US GPUs — and it's running on Huawei's homegrown Ascend chips. In this episode, we break down:
🔹 Longcat — the massive model built by Meituan, China's super app giant 🔹 China's exploding AI competitor ecosystem 🔹 Inside the **Huawei Ascend GPU architecture **: specs, costs, and energy tradeoffs vs. Nvidia 🔹 OpenAI's custom inference chip strategy 🔹 The software optimizations driving a 10,000x efficiency leap 🔹 GPT-5.6: Sol, Terra, and Luna models explained 🔹 New coding benchmarks with CursorBench 🔹 Meta's non-invasive brain-to-text research 🔹 Anthropic Science — AI built for researchers
https://novacut.ai/
Description:
Anthropic pulls access to Fable, and China responds the same day with GLM 5.2. In this episode we break down the escalating AI arms race, US export controls on chips and frontier models, and whether the "Great Firewall of America" is already here.
⏱️ Topics:
🔗 Links & Resources:
Fable
https://www.anthropic.com/news/claude-fable-5-mythos-5
https://support.claude.com/en/articles/14328960-identity-verification-on-claude
Midjourney
Full body ultrasound CT scanner
Xiaomi 1000tps
https://mimo.xiaomi.com/blog/mimo-tilert-1000tps (MiMo-V2.5-Pro-UltraSpeed: Pushing 1T-Parameter Model Generation Speed to 1000 TPS
Best opensource model
https://z.ai/blog/glm-5.2
📌 Timestamps in the chapters section above.
#AIPodcast #Anthropic #Fable #GLM52 #AIArmsRace #LLM #GenAI
0:00 Intro: Anthropic restricts Fable access
https://novacut.ai/
https://genaimeetup.com/
Anthropic has officially closed a $65 billion Series H at a $965 billion valuation, nearly 2.5x its valuation from just 100 days ago. Meanwhile, funding is flowing across the ecosystem: Frameworks AI at $15B, Baseten at $11B, OpenRouter's $113M Series B, and Cognition AI's $1B Series D.
NVIDIA went on an open-source super week with Nemotron 3 Ultra, Cosmos 3, and Nemotron 3.5 ASR. Microsoft dropped 5 new MAI models. Google released Gemma 4 12B, and Anthropic shipped Opus 4.8.
On the benchmarks front, DeepSWE crowns GPT-5.5 as the leader in long-horizon coding tasks, while ITBench shows even frontier models struggle with real-world SRE incidents — Claude Opus 4.7 tops out at just 47%.
Plus: Cloudflare acquires VoidZero to build the future of AI-native edge development, and Google is paying SpaceX $920M/month for compute.
Topics covered: • Anthropic's $65B Series H and path to $1T • Fireworks AI, Baseten, OpenRouter & Cognition funding rounds • Microsoft's 5 new MAI models • NVIDIA's open-source super week (Nemotron, Cosmos 3) • MiniMax M3, Gemma 4 12B, JetBrains Mellum2, Opus 4.8 • DeepSWE benchmark: GPT-5.5 leads long-horizon coding • ITBench: Frontier models under 50% on real SRE tasks • Cloudflare + VoidZero for AI-native edge dev • Google's $920M/month SpaceX compute deal
#AI #Anthropic #NVIDIA #OpenAI #AInews #TechNews #LLM
Anthropic formally confirmed the closure of its $65 billion Series H funding round at a post-money valuation of $965 billion. This represents a 2.5-fold increase over its $380 billion Series G valuation from February 2026, adding $585 billion in value in approximately 100 days
https://www.anthropic.com/news/series-h
Frameworks AI raising at 15B valuation representing a near fourfold increase from its $4 billion Series C valuation recorded in October 2025
processing 15 trillion tokens daily for major production clients including Cursor, Notion, and Perplexity
https://finance.yahoo.com/sectors/technology/articles/fireworks-ai-eyes-15-billion-174609357.html
Baseten is raising 1B at 11B valuation
annualized revenue, which skyrocketed from $200 million to $600 million over a single quarter
https://techstartups.com/2026/05/26/ai-inference-startup-baseten-in-talks-to-raise-1-billion-at-11-billion-valuation/
OpenRouter has secured a $113 million Series B funding
OpenRouter has experienced exponential traffic growth, with weekly production throughput expanding fivefold from 5 trillion to 25 trillion tokens over a six-month horizon
https://www.businesswire.com/news/home/20260526953416/en/OpenRouter-Raises-%24113-Million-CapitalG-led-Series-B-as-Weekly-Volume-Explodes-to-25T-Tokens
Further up the stack: Cognition AI secured a $1 billion Series D round led by Lux Capital and 8VC
https://cognition.ai/blog/series-d
MAI models:
https://microsoft.ai/models/
https://www.peoplematters.in/news/ai-and-emerging-tech/uber-imposes-dollar1500-monthly-ai-spending-limit-on-employees-amid-rising-costs-50073
Nvidia has executed an "Open-Source Super Week," positioning itself as a dominant software and model publisher:
https://www.minimax.io/models/text/m3
https://blog.google/innovation-and-ai/technology/developers-tools/introducing-gemma-4-12b/
https://www.jetbrains.com/mellum/
Opus 4.8
Benchmarks:
https://research.ibm.com/publications/developing-ai-agents-for-it-automation-tasks-with-itbench
ITBench-AA, an evaluation framework focusing on live Kubernetes incident response and Site Reliability Engineering (SRE) operations. Comprising 59 live, containerized SRE incident snapshots, the results are remarkably sobering: every frontier model scored under 50% on successful incident resolution, with Claude Opus 4.7 leading at 47% and GPT-5.5 following closely at 46%.
https://www.cloudflare.com/press/press-releases/2026/cloudflare-acquires-voidzero-to-build-the-future-of-the-ai-native-web/
This week on AI Meta, we break down Andrej Karpathy’s move to Anthropic, Claude’s growing developer
Plus: SpaceX IPO speculation, Cursor, Grok, and why the AI economy increasingly looks like a global
https://novacut.ai
https://novacut.ai/
Shashank and Mark break down a packed few weeks in AI: new open-source and local models, MCP and tool
They debate whether AI spending keeps compounding as models get cheaper and demand rises, or whether
In this episode, recorded across two continents with
From the publisher's feed

685 Listeners