Google Developers2.59 млн
Опубликовано 27 августа 2025, 1:00
Join host Logan Kilpatrick in discussion with some of the minds behind Google's new state-of-the-art image model, Gemini 2.5 Flash. Product and research leads from the Gemini team break down the technology behind its key capabilities, including interleaved generation for complex edits and new approaches to achieving character consistency and pixel-perfect control. With Nicole Brichtova, Kaushik Shivakumar, Mostafa Dehghani and Robert Riachi.
Listen to this podcast:
Apple Podcasts → goo.gle/3Bm7QzQ
Spotify → goo.gle/3ZL3ADl
Chapters:
0:37 - New model introduction
1:21 -Demo: Image editing
3:44 - Text rendering capabilities
4:44 Beyond human preference evals
6:44 - Text rendering as a proxy for quality
8:38 - Positive transfer between modalities
11:25 - Demo: multi-turn, context aware image generation
13:54 - Pixel-perfect editing and character consistency
15:51 - Interleaved image generation
17:59 - Specialized vs. native models
19:52 - Understanding nuanced prompts
20:59 - User feedback shaping model development
22:37 - Improvements in character consistency
24:17 - More natural looking images from team collaboration
26:41 - What’s next for image generation models
Watch more Release Notes → goo.gle/4njokfg
Subscribe to Google for Developers → goo.gle/developers
Speakers: Logan Kilpatrick, Nicole Brichtova, Kaushik Shivakumar, Mostafa Dehghani, Robert Riachi
Products Mentioned: Google AI, Gemini
Listen to this podcast:
Apple Podcasts → goo.gle/3Bm7QzQ
Spotify → goo.gle/3ZL3ADl
Chapters:
0:37 - New model introduction
1:21 -Demo: Image editing
3:44 - Text rendering capabilities
4:44 Beyond human preference evals
6:44 - Text rendering as a proxy for quality
8:38 - Positive transfer between modalities
11:25 - Demo: multi-turn, context aware image generation
13:54 - Pixel-perfect editing and character consistency
15:51 - Interleaved image generation
17:59 - Specialized vs. native models
19:52 - Understanding nuanced prompts
20:59 - User feedback shaping model development
22:37 - Improvements in character consistency
24:17 - More natural looking images from team collaboration
26:41 - What’s next for image generation models
Watch more Release Notes → goo.gle/4njokfg
Subscribe to Google for Developers → goo.gle/developers
Speakers: Logan Kilpatrick, Nicole Brichtova, Kaushik Shivakumar, Mostafa Dehghani, Robert Riachi
Products Mentioned: Google AI, Gemini
Свежие видео
Случайные видео























