Models
September 16, 2026

Google launches Gemini 3.8 Live and 3.5 Transcribe for real-time voice AI

Google has launched Gemini 3.8 Live and 3.8 Live Extended Thinking for real-time voice agents, alongside Gemini 3.5 Transcribe for low-latency speech recognition across more than 85 languages.

Google has introduced new Gemini Audio models for developers building real-time voice applications. Gemini 3.8 Live supports speech-to-speech conversations while executing tasks through asynchronous function calls, processing visual context, and maintaining dialogue across more than 97 languages.

Gemini 3.8 Live Extended Thinking adds configurable reasoning for complex, multi-step voice tasks. Google also highlighted Gemini 3.5 Transcribe, a dedicated speech-to-text model supporting more than 85 languages, automatic code-switching, custom vocabulary, and structured transcripts.

The Live models are available through the Gemini Live API, while developers can access the broader audio suite through Gemini API and Google AI Studio.

#
Google

Read Our Content

See All Blogs
LLM Models

LLM testing of Muse Spark 1.3

Sarankumar S

September 11, 2026
Read more
LLM Models

Harness engineering for AI agents: The missing layer for production deployment

Sarankumar S

September 11, 2026
Read more