The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Inside Nano Banana šŸŒ and the Future of Vision-Language Models with Oliver Wang - #748


Listen Later

Today, we’re joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Image—better known by its code name, ā€œNano Banana.ā€ We dive into the development and capabilities of this newly released frontier vision-language model, beginning with the broader shift from specialized image generators to general-purpose multimodal agents that can use both visual and textual data for a variety of tasks. Oliver explains how Nano Banana can generate and iteratively edit images while maintaining consistency, and how its integration with Gemini’s world knowledge expands creative and practical use cases. We discuss the tension between aesthetics and accuracy, the relative maturity of image models compared to text-based LLMs, and scaling as a driver of progress. Oliver also shares surprising emergent behaviors, the challenges of evaluating vision-language models, and the risks of training on AI-generated data. Finally, we look ahead to interactive world models and VLMs that may one day ā€œthinkā€ and ā€œreasonā€ in images.


The complete show notes for this episode can be found at https://twimlai.com/go/748.

...more
View all episodesView all episodes
Download on the App Store

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)By Sam Charrington

  • 4.7
  • 4.7
  • 4.7
  • 4.7
  • 4.7

4.7

419 ratings


More shows like The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

View all
Data Skeptic by Kyle Polich

Data Skeptic

479 Listeners

a16z Podcast by Andreessen Horowitz

a16z Podcast

1,094 Listeners

The AI in Business Podcast by Daniel Faggella

The AI in Business Podcast

169 Listeners

Super Data Science: ML & AI Podcast with Jon Krohn by Jon Krohn

Super Data Science: ML & AI Podcast with Jon Krohn

304 Listeners

NVIDIA AI Podcast by NVIDIA

NVIDIA AI Podcast

336 Listeners

Practical AI by Practical AI LLC

Practical AI

212 Listeners

Google DeepMind: The Podcast by Hannah Fry

Google DeepMind: The Podcast

197 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

91 Listeners

Dwarkesh Podcast by Dwarkesh Patel

Dwarkesh Podcast

498 Listeners

No Priors: Artificial Intelligence | Technology | Startups by Conviction

No Priors: Artificial Intelligence | Technology | Startups

134 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

210 Listeners

The AI Daily Brief: Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief: Artificial Intelligence News and Analysis

595 Listeners

Practical: AI & Business News by Practical News

Practical: AI & Business News

26 Listeners

AI + a16z by a16z

AI + a16z

34 Listeners

Training Data by Sequoia Capital

Training Data

37 Listeners