Artificial Intelligence

Google Launches Gemini 3.5 Live Translate for Real-Time Speech AI

MOUNTAIN VIEW, Calif.: Google DeepMind launches Gemini 3.5 Live Translate, a new audio model that delivers real-time speech-to-speech translation across more than 70 languages.

The model automatically detects which language someone is speaking without manual setup. Translated speech preserves the original speaker’s intonation, pacing, and pitch. The system supports more than 2,000 language pairings at launch globally.

✨ AI Summary

Google DeepMind has launched Gemini 3.5 Live Translate, an audio model providing real-time speech-to-speech translation across 70 languages. The system continuously streams audio while preserving the speaker’s original intonation and pitch. It features automatic language detection and operates with low latency, even in noisy environments.

Also read: Google Launches Two New AI Research Agents via Gemini API

The model is currently rolling out across Google Translate, Google Meet, and developer platforms. All generated audio includes SynthID watermarking to ensure transparency and address deepfake concerns. This release expands Google’s translation capabilities for enterprise and consumer users, providing accessibility through various hardware and software applications.

Traditional translation tools wait for a complete sentence before generating output to users. Gemini 3.5 Live Translate streams translated speech continuously while the speaker is talking. The model stays only a few seconds behind the original audio throughout sessions.

Also read: Google Launches Gemini-Powered Rambler Dictation Across Gboard App

The architecture handles noisy environments without degraded performance under stress conditions. Multilingual inputs work without manual switching between language tracks during conversations. Auto-detection eliminates the friction of selecting languages before each interaction begins.

The rollout spans multiple Google products immediately upon today’s announcement. Google Translate adds the model on Android and iOS apps simultaneously. The Gemini Live API and Google AI Studio open public preview access for developers globally.

Google Meet receives the model in private preview for enterprise customers this month. Android phones gain a new listening mode that delivers translations through the device earpiece. The hardware-agnostic approach means no headphones or dedicated devices are necessary.

SynthID watermarking marks all AI-generated audio output from the system automatically. The move addresses growing concerns around deepfake voice content across consumer platforms. Google positions the watermarking as essential transparency for live translation use cases.

“Our latest audio model takes real-time speech translation to the next level by delivering low-latency translation,” says Google in the launch announcement.

The launch intensifies pressure on Microsoft Translator, DeepL, and Meta’s translation efforts across enterprise and consumer markets. Google previously dominated translation research for over two decades.

Amita Parul

Amita Parul is an Independent journalist with experience in reporting and commentary on current events and sociopolitical developments. She contributes original reporting and analysis that aligns with Tea4Tech’s editorial standards for accuracy, transparency, and context, focusing on business and technology trends. || Amita covers emerging news stories and provides explanatory insights that help readers understand both the events and their implications.

Recent Posts

OpenAI Launches GPT-6 Astra With Advanced Computer and AI Capabilities

New Delhi: OpenAI has launched GPT-6 Astra, its latest and most advanced AI model, with…

4 days ago

India’s Rilo Joins Adobe in AI Marketing Automation Deal

New Delhi: Adobe has acquired Rilo, an India based marketing intelligence startup, in a deal…

1 week ago

Anthropic Levels Up Claude Fable 5.1 With More Power, Less Cost

New York: Anthropic has launched Claude Fable 5.1 and Claude Mythos 5.1, its latest AI…

1 week ago

Perplexity Launches ‘Portable Computer’, a Local AI Agent Built With Nvidia

Bengaluru: Perplexity AI has launched a new product called Portable Computer, an AI agent that…

2 weeks ago

India’s Ringg Bags $10M From Peak XV for Voice AI Push

New Delhi: Bengaluru-based voice AI startup Ringg has raised $10 million from Peak XV Partners,…

2 weeks ago

Google Halo Explained: A New Way to See What Your AI Agent Is Doing

New York: Google is moving forward with Android Halo, a new system-level feature designed to…

2 weeks ago