Google launched Gemini 3.8 Live. NVIDIA dropped new research on NemotronLabs VoiceChat. Meta expanded business calling through Muse.
3 of the biggest voice AI updates this week. Full roundp 👇
• Google: Gemini 3.8 Live can keep responding while a tool finishes its work. Devs can adjust reasoning through the Extended Thinking version.
• NVIDIA: NemotronLabs VoiceChat combines full-duplex speech with tool calling and reports 82.5% tool-selection F1 on Full-Duplex-Bench 3.0.
• Nuance Labs: $50M Series A led by Lightspeed to fund a model that handles audiovisual input and output together. Research preview expected later this year.
• Meta Muse: US business calls are rolling out to more beta users. People who asked for it earlier get priority.
• Instinct: Concierge is reaching early users with tasks like booking reservations by phone, requesting dentist cancellation slots and sorting out cable billing issues.
• Speechmatics: Agent STT uses the Linden 1 model to better catch names, account numbers and short replies.
• Deepgram: India endpoint is now GA. The new Nova-3 Pharma model focuses on drug names and pharma terms.
• SpaceXAI: Grok Voice Transcribe 2.0 does batch transcription at $0.10 per audio hour and streaming at $0.20.
• AssemblyAI: The new Dictation API turns speech into cleaned-up text, with Blurt as an open-source dictation app.
• DeepL: Speakers keep their own voice when translating across 12 languages. Meeting translation is out of beta on desktop. External browser access is still in testing.
• Pindrop: BotStopper can now be bought on its own to flag automated callers. Its AI Voice Consortium registry is past 5,000 voices.
• Retell: Conductor now works across the whole dashboard and can dig into calls using info from your workspace. Users approve any production edits it proposes.
• ServiceNow / Krisp: ServiceNow is using Krisp's VIVA SDK for voice agents handling IT, HR and customer service calls.
•
smallest.ai: Bolna (YC F25) users can now pick Pulse for STT and Lightning for TTS.
• LiveKit: Eligible startups get cloud and inference credits plus waived plan fees worth ~$23K, across 7 inference providers.
•Cekura: We launched our latest STT benchmarks, check them out at
benchmarks.cekura.ai/stt
The use cases are getting really specific, like recognizing medication names or calling a restaurant to book a table.
At
@cekuraAi we test how voice agents handle what goes wrong along the way. Every new capability means new things to test, like a misheard prescription or a booking that needs human approval.