TWIML AI Podcast
Inside Nano Banana 🍌 and the Future of Vision-Language Models with Oliver Wang - #748
Sep 23, 2025
· 1h 3m
Today, we’re joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Image—better known by its code name, “Nano Banana.” We dive into the development and capabilities of this newly released frontier vision-language model, beginning with the broader shift from specialized image generators to general-purpose multimodal agents that can use both visual and textual data for a variety of tasks. Oliver explains how Nano Banana can generate and iteratively edit images while maintaining consistency, …
이 에피소드는 아직 녹음되지 않았습니다
STT.ai을 사용하여 AI로 이 에피소드를 기록합니다. 발음기 감지, 타임스탬프, 다양한 형식으로 내보내기를 통해 정확한 텍스트를 얻으십시오.
스피커 감지
단어 수준 시간 스탬프
SRT, TXT, JSON으로 내보내기