
Oliver Wang
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang
- Published
- September 23, 2025
- Duration
- 1h 3m
- Summary source
- description
- Last updated
- Jul 5, 2026
Discusses google-ai, multimodal.
Summary
Today, we’re joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Image—better known by its code name, “Nano Banana.” We dive into the development and capabilities of this newly released frontier vision-language model, beginning with the broader shift from specialized image generators to general-purpose multimod…
Intelligent Report
Sign in to read teasers, or upgrade to Research Pro to commission intelligent report for this episode. Learn more →
Show notes
Today, we’re joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Image—better known by its code name, “Nano Banana.” We dive into the development and capabilities of this newly released frontier vision-language model, beginning with the broader shift from specialized image generators to general-purpose multimodal agents that can use both visual and textual data for a variety of tasks. Oliver explains how Nano Banana can generate and iteratively edit i
Themes
- google-ai
- multimodal