Sponsored by Byond Boundrys Consulting - Empowering Ideas, Delivering Results
Google New Intermediate

Gemini 3.5 Live Translate

Gemini 3.5 Live Translate is a verified current AI model with official specs, pricing or access details, capabilities, best use cases and key limitations.

Realtime Translation ModelTextAudio Freemium
In plain English

What is this model and why does it matter?

Gemini 3.5 Live Translate is Google's low-latency speech-to-speech model for bidirectional realtime translation with natural audio output and transcript text.

Live translationTravel appsMultilingual meetingsContact centersRealtime interpreter products
Model overview

Gemini 3.5 Live Translate: features, use cases and important details

Gemini 3.5 Live Translate is a current AI model verified from first-party Google sources.

Gemini 3.5 Live Translate verified specifications

Gemini 3.5 Live Translate is Google’s low-latency speech-to-speech model for bidirectional realtime translation with natural audio output and transcript text. Its verified context or usage limit is 131072 input tokens, with 65536 output tokens maximum output.

Gemini 3.5 Live Translate pricing and access

Paid audio pricing is $3.50/M input audio tokens and $21/M output audio tokens, equivalent to about $0.0368 per translated audio minute overall.

Gemini 3.5 Live Translate best uses

Live translation, Travel apps, Multilingual meetings, Contact centers, Realtime interpreter products.

Gemini 3.5 Live Translate limitations

No batch API, No grounding or tools, Preview availability can change.

Gemini 3.5 Live Translate capabilities and use cases

In addition, its main capabilities include Realtime speech translation, 70+ languages, Audio-to-audio, Transcript output and Low-latency Live API. For example, common use cases include Interpreters, Meetings, Travel, Customer support and Voice apps.

Who should consider Gemini 3.5 Live Translate?

In practice, this model may suit Live translation, Travel apps, Multilingual meetings, Contact centers and Realtime interpreter products. Also, notable strengths include 70+ languages, Natural speech output, Free tier and Low-latency design. However, review trade-offs such as No batch API, No grounding or tools and Preview availability can change before adopting it.

Gemini 3.5 Live Translate pricing and access

Meanwhile, Paid audio pricing is $3.50/M input audio tokens and $21/M output audio tokens, equivalent to about $0.0368 per translated audio minute overall. Effective combined audio cost is approximately $0.0368 per minute at the documented token rate.

Official resources and verification

Use the official model website, official documentation, pricing or release source and additional primary source to confirm current availability, limits and pricing. Product details can change after publication, so rely on primary documentation for final decisions.

Compare with other AI models

Next, continue your research in the AI models directory, Google models and Realtime Translation Model models. Compare providers, pricing, modalities and practical limitations side by side to choose the right model for your workflow.

Get started

How to use this model

  1. Create a Gemini API key.
  2. Connect through the Live API.
  3. Call gemini-3.5-live-translate-preview.
  4. Stream spoken audio.
  5. Receive translated audio and transcript text.
Copy and try

Example prompts

  • Translate this English speech to Hindi in realtime.
  • Interpret this bilingual meeting naturally.
  • Translate the caller while preserving conversational tone.
Capabilities

What it can do

  • Realtime speech translation
  • 70+ languages
  • Audio-to-audio
  • Transcript output
  • Low-latency Live API
Best for

Practical use cases

  • Interpreters
  • Meetings
  • Travel
  • Customer support
  • Voice apps
Pricing

What does it cost?

Paid audio pricing is $3.50/M input audio tokens and $21/M output audio tokens, equivalent to about $0.0368 per translated audio minute overall.

Input$3.50 / 1M audio tokens (~$0.0053/min input)
Output$21 / 1M audio tokens (~$0.0315/min output)
Simple summaryEffective combined audio cost is approximately $0.0368 per minute at the documented token rate.

What stands out

  • 70+ languages
  • Natural speech output
  • Free tier
  • Low-latency design

Things to consider

  • Preview status
  • No function calling
  • Audio-only input
Limitations

Important restrictions and trade-offs

  • No batch API
  • No grounding or tools
  • Preview availability can change
SimplifyAITools verdict

Our editorial take

A very current model with clear user interest around live multilingual voice translation and realtime AI applications.

References

Primary sources

  1. Open source 1 ↗
  2. Open source 2 ↗
  3. Open source 3 ↗