Live dialogue
Fluid and natural live dialogue and translation capabilities, for powerful voice-first applications
Fluid and natural live dialogue and translation capabilities, for powerful voice-first applications
Create agents capable of handling complex tasks and using tools, while engaging in natural conversations.
Best for complex reasoning and high-complexity tasks. Thinks and narrates its progress in real time to solve difficult, multi-step challenges.
Best for efficient high-volume deployments. Delivers levels of reasoning required for custom voice agents while keeping costs low.
Best for near real-time speech-to-speech translation. Overcomes language barriers across 70+ languages while maintaining the speaker’s natural tone and rhythm.
Build natural-sounding and reliable voice agents with our live dialogue models.
Handles background tasks while speaking, using deep reasoning to solve complex queries on the fly – without stalling the conversation.
Understands visual input to grasp deeper context during conversations. This allows for richer, more natural, and more relevant responses.
Knows the difference between direct engagement and background chatter. Understands the rhythm of speech, knowing precisely when to talk – and when to stay silent.
Handles complex alphanumeric data – like phone numbers, order IDs, and email addresses– reliably and accurately.
See how you can use 3.8 Live Extended Thinking to build voice agents.
Overcomes language barriers by using Gemini’s speech-to-speech translation capabilities.
Delivers fluid speech-to-speech translation across 70+ languages and 2,000 language pairs.
Preserves the speaker’s original intonation, pacing and pitch to capture not just what they said, but how they said it.
Translates multiple languages in a single session – no need to change the settings.
Identifies the language being spoken and begins translation, without being told what it is.
Minimizes processing lag to eliminate awkward pauses – and keep conversations flowing naturally.
Gemini 3.8 Live & 3.8 Live Extended Thinking introduces major leaps over our predecessor, Gemini 3.1
| Name | 3.8 Live Extended Thinking | 3.8 Live | 3.5 Live Translate |
|---|---|---|---|
| Status | General Availability | General Availability | General Availability |
| Input |
|
|
|
| Output |
|
|
|
| Input tokens | 128k | 128k | 128k |
| Output tokens | 64k | 64k | 64k |
| Availability |
|
|
|
| Documentation | View developer docs | View developer docs | View developer docs |
| Model card | View model card | View model card | View model card |
The fastest path from prompt to production
Get started with cutting-edge AI models
Low-latency, real-time voice and video interactions with Gemini
Build, scale, and govern agents
Deploy specialized agents for product discovery, shopping, and customer service
Understand your world and communicate across languages
Collaborate, create, and communicate all in one place