Inside Nano Banana đ and the Future of Vision-Language Models with Oliver Wang - #748
Today, weâre joined by Oliver Wang, principal scientist at Google DeepMind and tech lead for Gemini 2.5 Flash Imageâbetter known by its code name, âNano Banana.â We dive into the development and capabilities of this newly released frontier vision-language model, beginning with the broader shift from specialized image generators to general-purpose multimodal agents that can use both visual and textual data for a variety of tasks. Oliver explains how Nano Banana can generate and iteratively edit images while maintaining consistency, âĻ
āĻāĻ āĻ āϧā§āϝāĻžā§ āĻāϤāĻŋā§āĻžāĻ āĻ āύā§āϞāĻŋāĻĒāĻŋ āĻā§°āĻž āĻšā§ā§ąāĻž āύāĻžāĻ
AI ā§° āϏā§āϤ⧠āĻāĻ āĻ āϧā§āϝāĻžā§ā§° āĻ āύā§āϞāĻŋāĻĒāĻŋ āĻā§°āĻŋāĻŦāϞ⧠STT.ai āĻŦā§āĻ¯ā§ąāĻšāĻžā§° āĻā§°āĻāĨ¤ āϏā§āĻĒāĻŋāĻāĻžāϰ āĻāĻŋāύāĻžāĻā§āϤāĻā§°āĻŖ, āϏāĻŽā§āĻāĻŋāĻšā§āύ āĻā§°ā§ āĻŦāĻšā§āĻŦāĻŋāϧ āĻŦāĻŋāύā§āϝāĻžāϏāϤ āĻāĻā§āϏāĻĒā§ā§°ā§āĻā§° āϏā§āϤ⧠āϏāĻ āĻŋāĻ āĻā§āĻā§āϏāĻ āĻĒā§ā§°āĻžāĻĒā§āϤ āĻā§°āĻāĨ¤