Gemini 3.8 Flash is here⚡️
it's our most intelligent workhorse model, delivering significant improvements from 3.7 Flash across software engineering, agentic tasks, and critical, multi-step reasoning in specialized domains
available at the same introductory price as 3.7 Flash at $0.75 per million input tokens and $3.75 per million output tokens
try it today via the Gemini API and in AI Studio: ai.studio
introducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, our most expressive audio generation models yet
these models enable creators, developers, and enterprises to create richer, more expressive audio experiences
try them via the Gemini API and in AI Studio: aistudio.google.com/generate…
if you're in SF, join us next week (september 24th) for a first-of-its-kind Gemini Audio event
chat with members of the team, enjoy refreshments, and experience demos
RSVP here: rsvp.withgoogle.com/events/g…
🔈 today we're introducing two new live dialogue models
Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intelligence and multi-step reasoning.
these audio models are available to build with in AI Studio and via the Gemini API ai.studio
today we’re sharing the first preview of the Gemini API docs right in AI Studio
we built the experience from the ground up to give you better consolidation and convenience: docs right where you’re building
get started today: ai.studio/docs
Lyria 3.5, our best-sounding music generation model, is now available in AI Studio, via the Gemini API, and in the Gemini app
this model brings more expressive vocals and richer musical arrangements, allowing you to craft tracks with higher fidelity
try it today: ai.studio
introducing agentic video understanding with Gemini
instead of using static processing (ingesting media at a fixed frame rate), the model can now take an active, goal-directed role in determining what to watch, at what speed, and through which modality (frames, audio, or transcript), fetching only the moments and signals needed
this new approach cuts costs by up to 66% and reduces token consumption by up to 88% while boosting accuracy
available now via the Gemini API and in AI Studio: ai.studio