ChengRang

Gemini 3.8 Live

AI Chatbots Freemium
This page covers a version or sub-product of Gemini. View Gemini overview →

A near-real-time voice conversation model released by Google DeepMind on 2026/9/15, with two variants: the standard version and Live Extended Thinking. The latter reasons while it speaks and targets complex multi-step tasks, ranking first on the Speech to Speech Quality Index at 82.6, scoring 68.6% on τ-Voice and 97.7% on Big Bench Audio. It supports mid-conversation switching across 97 languages, user interruptions, background tool calls and asynchronous long-running tasks, and embeds SynthID watermarks in its audio.

GoogleLive VoiceVoice AgentMultimodalChat
Visit Gemini 3.8 Live

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are two near-real-time voice conversation models released by Google DeepMind on September 15, 2026, which the company calls its most advanced real-time conversational models to date. The two models launched the same day and share the same foundation, differing in the scenarios they target: Gemini 3.8 Live is built for scale and cost efficiency, combining conversational intelligence, fluent interaction and visual grounding; Gemini 3.8 Live Extended Thinking targets highly complex tasks, with stronger multi-step reasoning that lets it reason and speak at the same time.

The interaction design of Extended Thinking is the most interesting change in this generation. It confirms requests with advance verbal cues like "Let me check that…", then keeps narrating its progress by voice while pushing multi-step tasks forward in the background, turning the wait into a perceptible conversation instead of leaving the call in silence. Both models can execute tool calls and API requests in the background while continuing to talk with the user, and both let users interrupt at any time and change their mind on the spot.

On the capability side, both models can automatically detect and switch languages mid-conversation, covering 97 languages; all generated audio carries an imperceptible SynthID watermark. Developers access them through the Gemini Live API, with a /live entry point in Google AI Studio. On the enterprise side they are in private preview via Gemini Enterprise, while on the consumer side they have already arrived in Search Live, Gemini Live, and Workspace's Docs, Gmail and Keep.

Key Features

Use Cases

Pros

Pricing

Google has not yet disclosed specific pricing, describing it officially as highly competitive relative to other frontier models and emphasizing the cost efficiency of 3.8 Live. Developers can try it through the Gemini API and Google AI Studio (the /live entry point); Gemini Enterprise is in private preview; Search Live is available directly to general users, Extended Thinking is available in Workspace Docs for Google AI Pro and Ultra subscribers, and all Google AI subscribers can use it in Gmail and Keep. Refer to the official page for actual pricing.

Summary

The Gemini 3.8 Live family addresses the two things users complain about most in voice assistants: long waits and not daring to interrupt. The standard version makes real-time conversation cheap and reliable, while Extended Thinking lets complex tasks run to completion within a call. Among the official results, Extended Thinking ranks first on the Artificial Analysis Speech to Speech Quality Index with 82.6, plus 68.6% on τ-Voice and 97.7% on Big Bench Audio, with voice agent task completion the focus of this upgrade cycle. Products handling voice customer service, real-time guidance or multi-step transactions should prioritize testing this generation of the Live API; if you only need text conversation, a general-purpose model like Gemini 3.8 Flash remains more straightforward.

Version History

Category
AI Chatbots
Pricing
Freemium
Tags
Google · Live Voice · Voice Agent
Website

Related Tools