Gemini 3.8 Live
Default Live API model; thinking_level not supported on gemini-3.8-live (use gemini-3.8-live-extended-thinking for deeper reasoning; pricing page lists it at the same rates as gemini-3.8-live, checked 2026-09-29). Previous: gemini-3.1-flash-live-preview, gemini-2.5-flash-native-audio-preview-12-2025. WebSocket endpoint is the standard Live API URL, not re-read today.
- Context window
- 131,072 tokens
- Max output
- 65,536 tokens
- Input
- text, image, audio, video
- Output
- text, audio
- License
- proprietary
- Pricing
- input: $0.75 · output: $4.5 · audio input: $3 · audio output: $12 (per 1M tokens (text in $0.75, text out $4.50; audio in $3.00 = ~$0.005/min, audio out $12.00 = ~$0.018/min)) source
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| Gemini Live API (WebSocket) | gemini-3.8-live | wss://generativelanguage.googleapis.com/ws/google.ai.generativelanguage.v1beta.GenerativeService.BidiGenerateContent | docs |
| Gemini Live API (WebSocket) | gemini-3.8-live-extended-thinking | — | docs |
| Web app | — | gemini.google.com | — |
Notable capabilities (3)
- Real-time multilingual voice agents: Low-latency speech-to-speech with near-real-time visual grounding; 97 languages with mid-conversation switching. source
- Asynchronous tool use while talking: Keeps the conversation going while tools run in the background, narrating progress ('Let me check that...'). source
- Extended Thinking variant tops S2S quality: gemini-3.8-live-extended-thinking ranked #1 on Artificial Analysis Speech-to-Speech Quality Index (82.6) and 97.7% Big Bench Audio. source
Native-audio model for real-time voice and video agents via the Live API (WebSocket, bidirectional streaming).
from google import genai
client = genai.Client()
async with client.aio.live.connect(model="gemini-3.8-live",
config={"response_modalities": ["AUDIO"]}) as session:
await session.send_client_content(turns={"parts": [{"text": "Hi!"}]})
Sources: model page, pricing, launch blog.
Other Google DeepMind models
Gemini 3.8 Flash TTS · Gemini 3.8 Flash · Gemini 3.5 Transcribe (and Transcribe Live) · Lyria 3.5 · Gemini 3.5 Flash-Lite · Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image) · Gemini Omni Flash (Omni 1.1 Flash) · Gemini Embedding 2 · Gemma 4 · Nano Banana 2 (Gemini 3.1 Flash Image) · Nano Banana Pro (Gemini 3 Pro Image) · Gemini Robotics 2 · Gemini Robotics ER 2 · Gemini Robotics On-Device 2 · Gemini 3.5 Live Translate · Gemini 3.1 Pro · Veo 3.1 · Genie 3 · Lyria RealTime · Gemini 3.7 Flash · Gemini 3.6 Flash · Gemini 3.5 Flash · Gemini 3.1 Flash TTS (preview) · Gemini 3.1 Flash Live (preview) · Lyria 3 (Clip / Pro) · Lyria 2 · Gemini 2.5 Flash Native Audio (Live, preview) · Gemini 2.5 Flash-Lite · Gemini 2.5 Flash · Gemini 2.5 Pro · Gemini 2.5 Flash TTS / Pro TTS · Gemini 3.1 Flash-Lite · Nano Banana (Gemini 2.5 Flash Image) · Gemini Robotics-ER 1.5 / 1.6 · Imagen 4