Gemini 3.8 Live Extended Thinking

by Google

Gemini 3.8 Live Extended Thinking is the higher-reasoning member of Google's Gemini 3.8 Live audio family, announced on September 15, 2026 and built for high-complexity tasks that need multi-step reasoning during a live voice conversation. It takes text, images, audio and video as input and returns text and audio, with a 131,072-token input limit and 65,536-token output limit. The model reasons in the background while it keeps talking, using natural verbal cues such as "Let me check that" and narrating progress on multi-step tasks. Background reasoning is set with thinking levels of low, medium or high, and function calling is asynchronous only, so tools keep running while audio continues to stream; search grounding is also supported. Google says it handles 97 languages, including switching language mid-conversation. On Google's numbers it scores 82.6 on Artificial Analysis' Speech to Speech Quality Index (the top overall rank at launch), 68.6% on the tau-Voice agentic task benchmark, 35.1% on Sierra's tau-Voice banking benchmark and 97.7% on Big Bench Audio. It fits voice agents and assistants that must complete complex, tool-heavy workflows without breaking conversational flow.

Call Gemini 3.8 Live Extended Thinking on Opper with the OpenAI SDK. Sign up without a credit card. Get started

Key info

Input
Output
Features
Context window
131K
Max output
66K
Input price
$0.75 /1M
Output price
$4.50 /1M
  • US residency available
  • Zero data retention with abuse monitoring
  • No training by default
  • GDPR DPA available

Available routes

Gemini 3.8 Live Extended Thinking runs on 1 route through the Opper gateway. Compare residency, zero data retention and training posture at a glance, with full data-handling detail per route below.

ProviderRegionZero data retentionTrainingInputOutputCache read
USNo$0.75$4.50Input price

Live routes and prices as JSON: api.opper.ai/v3/models?q=gemini%2Fgemini-3.8-live-extended-thinking

Uptime and availability

Gemini 3.8 Live Extended Thinking runs on a single route through the Opper gateway today, so its availability is that provider's availability.

100%route uptime, last 30 days
Single route, no failover target within this model.

Name a second model in the same request and the gateway tries it on retriable errors, so a busy hour never has to reach your users. Set up a fallback chain.

Measured over the last 30 days from each provider's official status feed via StatusGator. Refreshed hourly. See uptime for every provider Opper monitors.

Data handling per route

Each route hosting Gemini 3.8 Live Extended Thinking has its own privacy posture, residency, and GDPR terms. Postures are maintained by Opper with a last-verification timestamp.

Google Cloud — United States🇺🇸

Zero data retention with abuse monitoring: prompts and outputs are kept for abuse monitoring only, for a window the provider does not state, and are not used for training. No training on customer data. US; SCCs; DPA available.

Zero data retention
With abuse monitoring, window not stated.
Training
No training on customer data.
Logging
Abuse monitoring
Abuse monitoring
On by default, holds flagged content, 55 days, a person can view samples
GDPR DPA
DPA available
Transfer mechanism
SCCs

Get started

Call Gemini 3.8 Live Extended Thinking through the Opper gateway with one API key. Opper is drop-in compatible with the OpenAI, Anthropic and Google AI SDKs.

Gemini 3.8 Live Extended Thinking is a premium model on Opper. Sign up needs no credit card: you get an API key straight away and the free models work in the playground and the API. Add a card to use premium models, pay-as-you-go with no minimum.

Set it up with your agent

Copy this and paste it into a coding agent like Claude Code, Cursor or Codex and it'll wire up Opper for you.

Or call it directly

import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.OPPER_API_KEY,
baseURL: "https://api.opper.ai/v3/compat",
});
const completion = await client.chat.completions.create({
model: "gemini/gemini-3.8-live-extended-thinking",
messages: [{ role: "user", content: "Hello" }],
});
console.log(completion.choices[0].message.content);

Use Gemini 3.8 Live Extended Thinking in your apps

Pick your tool to see how to run it on Gemini 3.8 Live Extended Thinking through Opper.

Compare Gemini 3.8 Live Extended Thinking with…

Side-by-side on privacy, EU hosting, pricing, and benchmarks.

Other models from Google

Start building with 700+ models

One API key for every major provider, up and running in minutes.

Get startedView Documentation