Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Gemini 3.8 Live Extended Thinking

Gemini 3.8 Live Extended Thinking is a voice-oriented variant in Google's Gemini model family, introduced alongside the base Gemini 3.8 Live on September 15, 2026. According to Google's own developer documentation hosted on ai.google.dev, the model is documented under the Gemini API as a distinct entry, signaling that it is treated as a separately addressable endpoint rather than a minor patch to the live audio model. The Jetstream announcement frames the Extended Thinking version as a counterpart tuned for more demanding reasoning, building on the conversational strengths of the live line while pushing toward multi-step processing that voice-only assistants historically struggled to support.

Practically, this variant is aimed at developers who want a single model that can both hold natural spoken dialogue and carry out the chain-of-thought work needed for harder queries, such as stepwise troubleshooting during a voice call or reasoning aloud over a spoken prompt. Its positioning next to Gemini 3.8 Live suggests a design intent of keeping latency and conversational fluency intact while allocating extra budget for internal deliberation, making it a better fit than the base live model for assistants that must reason rather than merely respond. For teams evaluating it, the key fit is voice-first applications where richer thinking is required without sacrificing the real-time feel of the Live family.

Vercel AI Gatewaygoogle/gemini-3.8-live-extended-thinkinggemini

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
google/gemini-3.8-live-extended-thinking
Release date
Sep 15, 2026
Last updated
Sep 15, 2026
Input modalities
Output modalities
Capabilities
Base catalog fields only

Cost

Input token cost
$0.75
Output token cost
$4.50

Limits

Output tokens
0 tokens
Context window
0 tokens

Latest news about Gemini 3.8 Live Extended Thinking

Vercel AI Gateway

Coverage

A September 24, 2026 Google Cloud blog post by Gemini Live Group Product Manager Fabien Blanc-paques confirms the general availability of Gemini 3.8 Live with Live Avatar in Gemini Enterprise and references "our announcement of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking last week," explicitly naming the Exte The post details capability areas that apply to the broader Gemini 3.8 Live family: video avatars with synchronized lip-syncing, native speech-to-speech dialogue with interruption recovery, background tool calling, support for 97 languages with automatic detection, and simultaneous live camera feeds plus screen shares.

Vercel AI Gateway

Coverage

Yahoo's tech coverage frames Google's September 15, 2026 launch of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking in a competitive context, noting the Extended Thinking model's 82.6 score on Artificial Analysis' Speech to Speech Quality Index narrowly edges OpenAI's GPT-Live-1 Astra at 81.5 and xAI's Grok Voice The article also reports that Google is partnering with Salesforce, Lumeris, and Genspark to develop enterprise use cases for the technology, and confirms both models are available through the Gemini API and Google AI Studio. Gemini 3.8 Live is rolling out in Search Live, while the Extended Thinking variant is rolling

Vercel AI Gateway

CoverageBenchmark

DataCamp published a September 17, 2026 technical breakdown covering both Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, noting that Google released both speech-to-speech models on September 15, 2026 as successors to Gemini 3.1 Flash Live. It reports that Extended Thinking tops the Artificial Analysis Speech-to The article describes Gemini 3.8 Live Extended Thinking as the higher-reasoning sibling that "reasons and calls tools in the background while it keeps talking," suitable for complex multi-step tasks, with both variants generally available in the Gemini API at the same price as their predecessor. Pricing is unchanged fr

Vercel AI Gateway

Coverage

9to5Google's same-day coverage reports Google's announcement of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking as its most advanced live dialogue models, succeeding the 3.1 Flash Live launch from March. The Extended Thinking variant is described as offering increased intelligence and multi-step reasoning, design The piece details that the Extended Thinking model "reasons and speaks simultaneously," using early verbal cues such as "Let me check that…" and live progress narration for multi-step background tasks, and reiterates Google's benchmark claims of 82.6 on the Speech to Speech Quality Index, 68.6% on τ-Voice, 35.1% on τ-V

Vercel AI Gateway

CoverageRelease Notes

MarkTechPost covers Google's September 15, 2026 release of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking as native speech-to-speech models designed for production-grade real-time voice agents. Both are hosted in the Gemini Live API and Google AI Studio with no open-weights or self-hosted option. The Extended Th The coverage highlights that Gemini 3.8 Live supports near-real-time visual processing, automatic switching across 97 languages mid-conversation, and asynchronous function calling that streams audio while tools execute in the background. It also reports that the base model placed second in the Speech Agent Arena, and n

Vercel AI Gateway

Coverage

Unite.AI reports Google's September 15, 2026 announcement of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, a pair of live dialogue models rolling out across the Gemini API, Google AI Studio, Gemini Enterprise, Search Live, Gemini Live, and Google Workspace. The models were introduced by Tom Ouyang (principal e Gemini 3.8 Live is positioned for scale and cost efficiency with conversational intelligence, fluid dialogue, and visual grounding, while the Extended Thinking variant is aimed at high-complexity tasks requiring increased intelligence and multi-step reasoning. Unite.AI restates the Google-reported benchmark figures inc

Vercel AI Gateway

Official sourceOfficial

Google's official blog announces the launch of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking on September 15, 2026, described as the company's most advanced live dialogue models to date. The post is authored by Tom Ouyang (Principal Engineer) and Malini Jaganathan (Member of Technical Staff) on behalf of the Ge The launch covers availability through the Gemini API, Google Workspace, and the Gemini app, with features including real-time visual context, background tool execution, and mid-conversation language switching. Google reports benchmark results for the Extended Thinking model including an 82.6 score on Artificial Analys

Vercel AI Gateway

Official sourceRelease Notes

Google's Gemini API release notes for September 15, 2026 document the general availability of Gemini 3.8 Live (gemini-3.8-live) and Gemini 3.8 Live Extended Thinking (gemini-3.8-live-extended-thinking) as two new audio-to-audio models for real-time voice applications using the Live API. The base model is described as t The Extended Thinking variant is documented as a high-reasoning audio-to-audio model that supports background reasoning during live audio interactions, recommended when higher background reasoning is required. The changelog entry points developers to the Live API guide, the Capabilities guide, and the Thinking guide to

Videos about Gemini 3.8 Live Extended Thinking

More models around Gemini 3.8 Live Extended Thinking