Robots Atlas>ROBOTS ATLAS
Gemini 3.8 Live

Gemini 3.8 Live

3.8 Liveย ยทย Family: Gemini
Google's cost-efficient voice model โ€” near real-time conversation and visual grounding, 97 languages, SynthID audio watermarking.
โœ“ Activeโœ“ Public accessMultimodalAudioAudio๐Ÿ“ Gemini
Release date
15 September 2026
Access:APIHostedDeployment:โ˜ Cloud

Overview

Gemini 3.8 Live is Google DeepMind's cost-efficient voice model, combining conversational intelligence, fluid dialogue and visual grounding, built for scale. It processes voice and visual inputs in near real time, supports 97 languages with automatic mid-conversation language detection and switching, can call tools and APIs, and executes background tasks without interrupting the dialogue. All generated audio is watermarked with SynthID.

The model rolled out on September 15, 2026 โ€” for developers via the Gemini API and Google AI Studio, and for users in Search Live, among others. It is marketed as highly cost-effective and ranked second in the Speech Agent Arena. A higher-tier variant, Gemini 3.8 Live Extended Thinking, targets high-complexity tasks.

Classification
MultimodalAudioAudio
Family: Gemini
Access & deployment
APIHosted
Cloud
Weights: Closed
Key parameters
โœ“ Tools
๐Ÿ“ฅ Input: text, audio, image, video

Technical specification

License
Proprietary
Features:โœ“ Tool use
Modalities
โฌ‡ Input
textaudioimagevideo
โฌ† Output
textaudio

Benchmark results

1 benchmark
Speech Agent Arena
2. miejsce
๐Ÿ“… 15 Sept 2026๐Ÿ“„ Google DeepMind (blog)
Marketed as highly cost-effective; ranked second.