Behind the scenes of Google's state-of-the-art "nano-banana" image model

45 428
12.4
Следующее
Популярные
13 дней – 1 9103:45
Mo’s Story: Starting Families
Опубликовано 27 августа 2025, 1:00
Join host Logan Kilpatrick in discussion with some of the minds behind Google's new state-of-the-art image model, Gemini 2.5 Flash. Product and research leads from the Gemini team break down the technology behind its key capabilities, including interleaved generation for complex edits and new approaches to achieving character consistency and pixel-perfect control. With Nicole Brichtova, Kaushik Shivakumar, Mostafa Dehghani and Robert Riachi.

Listen to this podcast:
Apple Podcasts → goo.gle/3Bm7QzQ
Spotify → goo.gle/3ZL3ADl

Chapters:
0:37 - New model introduction
1:21 -Demo: Image editing
3:44 - Text rendering capabilities
4:44 Beyond human preference evals
6:44 - Text rendering as a proxy for quality
8:38 - Positive transfer between modalities
11:25 - Demo: multi-turn, context aware image generation
13:54 - Pixel-perfect editing and character consistency
15:51 - Interleaved image generation
17:59 - Specialized vs. native models
19:52 - Understanding nuanced prompts
20:59 - User feedback shaping model development
22:37 - Improvements in character consistency
24:17 - More natural looking images from team collaboration
26:41 - What’s next for image generation models

Watch more Release Notes → goo.gle/4njokfg
Subscribe to Google for Developers → goo.gle/developers

Speakers: Logan Kilpatrick, Nicole Brichtova, Kaushik Shivakumar, Mostafa Dehghani, Robert Riachi

Products Mentioned: Google AI, Gemini
Случайные видео
18.12.24 – 17 8421:13
Tested for Everyday Life | Samsung
11.10.21 – 16 5091:11
Happy 15th Birthday, Google Docs!
17.02.15 – 2 046 3896:14
Audio Technica ATH-M70X Review!
автотехномузыкадетское