Sr. Director & Product Lead, Gemini Models @GoogleDeepmind. @Stanford CS. Opinions my own.

Help us build the future of voice-based experiences -- come join us for Gemini Audio | At Night in SF on Sept 24! It'll be an evening dedicated to the next frontier of voice-first AI. Meet the product and research teams behind our latest Gemini Audio models, experience hands-on demos, and network with builders and founders. More info here:rsvp.withgoogle.com/events/g…
6
7
81
11,972
Voice is continuing to become a larger way of how we interact, and we're building models that enable the creativity, expressivity, and naturalness you need! Today, we released two new TTS models that allow you to build new, original, expressive voices, or use your own, to generate the highest quality creations across 100+ languages!!
3
3
51
2,451
And, a week ago, we released SOTA 3.8 Live and 3.8 Live Thinking models for you to build more intelligent voice agents at scale (check out this example of a coding partner!) You can use these models in the APIs, Gemini Enterprise, Gemini Live, Search Live, and Docs Live...and there's even more fun coming soon.
1
2
332
Tulsee Doshi retweeted
We’re introducing Gemini 3.8 Live and 3.8 Live Extended Thinking – our best conversational AI. The models talk, think, and handle tasks in the background without breaking your flow. 🧵
355
380
3,637
662,592
Tulsee Doshi retweeted
Gemini Omni Flash is getting more controls (extension is finally here!) and resolutions with our latest update - so you can quickly iterate at 360p and scale up to 4k 🎉
Gemini Omni 1.1 Flash is our newest multimodal model for video generation and editing. It delivers a new suite of creative capabilities and controls for developers 🎥 With this update you can: 🎬 Extend your scenes 🎯 Specify starting and ending frames of a shot ➕ Add video input references ✨ Upscale your favorite takes up to 4K ⚡ Test ideas quickly in 360p See these in action 🧵
5
4
31
2,063
Tulsee Doshi retweeted
Look what you can do: - Tell Gemini what you want done - Gemini gets to work This is the Year of Voice
Gemini Live is moving beyond conversation to handle complex tasks on your behalf. With new features in Live like Daily Brief, Gemini Spark, Personal Intelligence, and @Gmail inbox management, you can talk through your day and delegate your to-dos without missing a beat. 🧵
26
20
420
45,336
If you've been using our Gemini 3.5 Transcribe model in gboard on Android or in the Gemini Mac App, you'll know it's a game changer for getting things done with voice. We're now bringing that magic to the API! Voice will continue to become more of a primary way we engage with agents, and I'm excited for you to use 3.5 Transcribe to build!
Say hello to Gemini 3.5 Transcribe! - Build apps that understand user speech / intent, even w/ multiple speakers! - Auto-detection of 85+ languages out of the box - Custom vocab adaptation for specialized jargon... SGTM:) API available now in @GoogleAIStudio and Gemini Enterprise, or try it in the Gemini app on macOS or Rambler on Android! More details: blog.google/innovation-and-a…
12
4
126
14,245
Tulsee Doshi retweeted
Gemini 3.7 Flash smashed previous Gemini growth records in its first week, making it our fastest growing model yet. Great to see the huge excitement from our developer community! Now running in Search and @Geminiapp too.
Gemini 3.7 Flash from @Google on ARC-AGI (Verified): - ARC-AGI-2: 84.6%, $0.25/task - ARC-AGI-1: 95.5%, $0.12/task Gemini 3.7 Flash stands out for its low cost and high scores on ARC-AGI-1 and ARC-AGI-2 relative to other frontier models.
358
313
3,882
682,463
Tulsee Doshi retweeted
In the spirit of making Antigravity accessible everywhere, we just launched remote control from the app and headlessly from the terminal. Happy coding, the usage limits are quite generous 🙂
52
16
519
25,265
The Gemma family has officially surpassed 1 billion downloads! The response to Google’s Gemma family of open models has been nothing short of inspiring. From local on-device apps to orbit-based satellites, developers are proving what’s possible with small open models, and creating over 100,000 Gemma model variants. Thank you to every developer, researcher, and builder who has made Gemma a part of their journey -- it's been amazing to see our Gemmaverse ecosystem thriving!!
When we first introduced @GoogleGemma, our family of open models, our goal was to give developers the tools to build responsible, innovative AI applications anywhere. Today, Gemma models have surpassed a billion downloads. 🎉 Over the past two years, developers have published over 100,000 Gemma model variants and built a thriving innovation ecosystem we call the Gemmaverse. We’re taking a look at how the Gemmaverse is making an impact across the globe ↓🧵
3
5
63
5,577
Tulsee Doshi retweeted
3.7 Flash : )
303
151
2,739
445,967
Gemini 3.7 is here, only 3 weeks after 3.6 Flash, with major improvements at an introductory price that is half of 3.6 Flash. We've been using it internally as our daily driver, and it's just...better to use! Better planning, more reliable and adaptable, and not to mention more capable across a number of benchmarks :)
11
7
243
9,780
Build fun things with 3.7 Flash and let us know what you think!
2
15
1,139
Can’t wait for everyone to try Rambler — one of my favorite models! As someone who has wrist issues and often voice types, this feature is a game changer.
The new Pixel 11 line up is here, designed for Gemini Intelligence and with lots of new helpful features. A few I’m excited about: Rambler, a new voice input built for how you actually talk. Magic Capture for effortless photos. And HiLight, a subtle glow when your phone is face down that lets you know when important people in your life are trying to call you.
1
51
6,306
Tulsee Doshi retweeted
1B+ people are now using @Geminiapp every month to spark new ideas and get things done. It’s our fastest growing product ever, and our 14th to hit the 1B-user mark. Kudos to @JoshWoodward & the entire Gemini team, and thank you to everyone who has been on this journey with us - much more to come!
1,159
650
7,723
4,637,020
Tulsee Doshi retweeted
Lyria 3.5 in Google Flow Music ❤️❤️ Deepmind really cooked with this one 🔥🔥 really I didn't expected the lyrics could be this good Video and song generated by producer itself
Deepmind really cooked with this 🔥🔥 I will share two songs: English, Hindi soon....
6
6
50
5,830
Tulsee Doshi retweeted
For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making that dream a reality: Introducing Gemini Robotics 2 from @GoogleDeepMind, the intelligence layer powering the next generation of truly adaptable robots. This major advance unlocks intelligent whole-body control, advanced dexterity, and even multi-robot collaboration 🤯. Ok but... how does a robot actually "think"? Real-world tasks take time and planning. To manage that complexity, our new embodied reasoning model, Gemini Robotics ER 2, acts as the robot’s high-level brain, enhancing the robot’s capabilities to: — Observe the environment — Reason about the actions needed to complete the task — Coordinate with the vision-language-action model to carry out actions — Track progress until the job is done This setup allows robots to execute complex multi-step workflows, self-correct if a step fails, and adapt to completely novel situations. Learn more about Gemini Robotics ER 2 (and our two other brand new models) here: goo.gle/4x4E8q6
313
781
5,247
785,679
Excited to see Gemini 3.6 Flash deliver frontier-level accuracy at $2.41/query and 164s. Try Gemini 3.6 Flash for financial use cases, and share your feedback! A great combination of speed, cost-efficiency, and high performance
New results in: Gemini 3.6 Flash achieves 46.3% on FrontierFinance, our open benchmark for finance AI agents. Ahead of Claude Opus 4.8, on par with GPT-5.6 Sol, and a big jump over past Gemini generations. What stands out is the efficiency: it hits that score at just $2.41 per query, cheaper than both, and at 164s it's among the fastest models we've tested on this harness. In our analysis, it's notably stronger than Opus 4.8 at surfacing qualitative and contextual insight, and leads on Screening & Discovery, one of the benchmark's hardest use cases. More about FrontierFinance and the full benchmarking results 👇
61
22
498
323,011
Tulsee Doshi retweeted
Gemma model family crossed the 900M downloads. What a milestone!
46
27
335
690,358
Tulsee Doshi retweeted
A strong and secure open ecosystem is important for the world to benefit from AI. We’ve always supported and contributed heavily to open source and science from Jax to Transformers to AlphaFold to Gemma open models which have now been downloaded 300M+ times. And the standards framework we’ve proposed supports responsible deployment of both open and proprietary models.
For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-W…
262
774
7,427
1,164,259