Radio and PodcastRadio and PodcastLive Radio & Podcasts
Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 artwork
Technology

Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748

This Week in Machine Learning & Artificial Intelligence (AI) Podcast by TWIML

Sep 23, 202563:39Technology

Today, we’re joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Image—better known by its code name, “Nano Banana.” We dive into the development and capabilities of this newl...

About This Episode

Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 is an episode from This Week in Machine Learning & Artificial Intelligence (AI) Podcast by TWIML. Today, we’re joined by Oliver Wang, principal scientist...

Listen Online

Use the player on this page to stream the episode online.

Episode Details

Published Sep 23, 2025, 63:39 long, audio available.

Questions About This Episode

What is Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 about?

Today, we’re joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Image—better known by its code name, “Nano Banana.” We dive into the development and capabilities of this newly released frontier vision-language model, beginning with the broader shift from specialized image generators to general-purpose multimodal agents that can use both visual and textual data for a variety of tasks. Oliver explains how Nano Banana can generate and iteratively edit images while maintaining consistency, and how its integration with Gemini’s world knowledge expands creative and practical use cases. We discuss the tension between aesthetics and accuracy, the relative maturity of image models compared to text-based LLMs, and scaling as a driver of progress. Oliver also shares surprising emergent behaviors, the challenges of evaluating vision-language models, and the risks of training on AI-generated data. Finally, we look ahead to interactive world models and VLMs that may one day “think” and “reason” in images. The complete show notes for this episode can be found at

Where can I listen to Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748?

You can listen to Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 online on Radio and Podcast. Open the player on this page to stream the available audio.

Which podcast is Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 from?

Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 is an episode from This Week in Machine Learning & Artificial Intelligence (AI) Podcast by TWIML.

How long is this episode?

This episode is 63:39 long.

When was this episode published?

This episode was published on Sep 23, 2025.

Can I save Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 for later?

Yes. Use the heart button on the episode page to add it to your favorite episodes list.

Are there related episodes from This Week in Machine Learning & Artificial Intelligence (AI) Podcast?

Yes. This page shows related episodes from This Week in Machine Learning & Artificial Intelligence (AI) Podcast when more episodes are available from the podcast feed.

Quick Answers About This Episode

Where can I listen to Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748?

You can listen to Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 on this page when the episode audio is available from the podcast feed.

Which podcast is this episode from?

Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748 is from This Week in Machine Learning & Artificial Intelligence (AI) Podcast by TWIML.

What are the episode details?

Published Sep 23, 2025 and 63:39 long