Signum
Feed
Useful signal15 Sept 2026high confidence

Google launches Gemini 3.8 Live and 3.8 Live Extended Thinking voice models with real-time visual context and background task execution

Google DeepMind launched two new voice/audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, rolling out starting today via the Gemini API, Google AI Studio, Gemini Enterprise (private preview), Search Live, Gemini Live app, and (for Extended Thinking) Google Workspace (Docs, Gmail, Keep) for AI Pro/Ultra subscribers. Features include near real-time visual context processing, automatic mid-conversation switching among 97 languages, background tool/API execution without interrupting conversation, and simultaneous reasoning-while-speaking for complex multi-step tasks. Audio outputs are watermarked with SynthID.

CapabilityAccessAdoptionEconomics

Entities: Google DeepMind, Gemini 3.8 Live, Gemini 3.8 Live Extended Thinking, Google Workspace, Google AI Studio, Gemini Enterprise

69Useful signal
1 source
1 primary
Was this useful?
01

What happened

Google DeepMind released two voice/audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, rolling out from today across the Gemini API, Google AI Studio, Gemini Enterprise (private preview), Search Live, the Gemini Live app, and, for Extended Thinking, Workspace (Docs, Gmail, Keep) for AI Pro/Ultra subscribers. Claimed features include near real-time visual context, automatic switching across 97 languages mid-conversation, background tool execution without interrupting the conversation, and SynthID watermarking on audio output. Google cites third-party benchmark placements, including top spot on Artificial Analysis's Speech to Speech Quality Index (82.6) and a 97.7% score on Big Bench Audio. No pricing was disclosed despite repeated references to being "cost-effective."

02

Why it matters

This affects developers building voice agents (via API access), enterprises evaluating Workspace and Gemini Enterprise voice tools, and consumers using Search Live or the Gemini app. The benchmark numbers are independently checkable, which gives this more substance than a typical announcement, and if the SSQI and agentic task scores hold up under third-party testing, it strengthens Google's position against rivals like OpenAI's voice models and specialised players such as Sierra and LiveKit. But without pricing, nobody can yet judge whether it is commercially competitive, and the practical impact for most businesses will depend on how it performs outside controlled benchmarks.

03

What is noise

The "most advanced" framing and repeated "cost-effective" claims are vendor packaging with no pricing to back them up. This is an iterative point release in a voice-model line that Google has been updating every few months, not a structural shift, so language implying a breakthrough moment should be discounted. The long list of partner names (Sierra, ServiceNow, Agora, LangChain, Vercel, etc.) reads as ecosystem-marketing padding rather than evidence of deep integration or committed usage.

04

Watch next

  1. 01Whether Google publishes concrete API pricing for Gemini 3.8 Live and Extended Thinking within the next few weeks, and how it compares to OpenAI's and Anthropic's voice offerings
  2. 02Independent verification of the Artificial Analysis SSQI 82.6 and Big Bench Audio 97.7% scores by other benchmarking outlets or developers, not just Google's own citation
  3. 03Whether any of the named partners (Sierra, ServiceNow, Salesforce, Genspark, Lumeris) publish their own case studies or usage data confirming real production deployment within the next one to two quarters

Coverage

1 story

More capability signals

Full feed →