Gemini 3.5 Live Translate
Gemini 3.5 Live Translate is Google’s latest audio model for live speech-to-speech translation, delivering near real-time translation in more than 70 languages. The model automatically detects multilingual input and generates smooth, natural-sounding translated speech that preserves the speaker’s intonation, pacing, and pitch. Unlike turn-by-turn translation systems that wait for someone to finish speaking before responding, Gemini 3.5 Live Translate processes speech as it streams and generates translated audio continuously, balancing the need for context with the need to stay in sync. It stays only a few seconds behind the speaker throughout a session, helping conversations feel more fluid and natural, without awkward pauses. It is built for multilingual calls, meetings, lessons, broadcasts, live interpretation, dubbing, simultaneous translation, and voice translation applications.
Learn more
Google Cloud Translation API
Make your content and apps multilingual with fast, dynamic machine translation available in thousands of language pairs.
The basic edition of the Translation API translates the texts of your website and your applications into more than 100 languages instantly. The Advanced edition offers dynamic results just as quickly as the Basic edition, but also includes other customization features, which is very important when you use phrases or terms that are specific to specific areas and contexts. The pre-trained model of the Translation API supports over a hundred languages, from Afrikaans to Zulu. With AutoML Translation you can create custom models in more than fifty language pairs. Thanks to the Translation API glossary, the content you translate will remain true to your brand. You just have to indicate which vocabulary you want to give priority to and save the glossary file in your translation project.
Learn more
Azure Speech Translation
Translate audio from more than 30 languages and customize your translations for your organization’s specific terms, all in your preferred programming language. Benefit from fast, reliable speech translation powered by neural machine translation technology. Generate speech-to-speech and speech-to-text translations with a single API call. Speech Translation captures the context of full sentences to provide accurate, fluent translations and improve communication between speakers of different languages. Customize speech recognition and translation for terminology specific to your business or industry. Train and deploy a custom translation system, without requiring machine learning expertise. Speech Translation can remove verbal fillers ("um," "uh," and coughs) and repeated words, add proper punctuation and capitalization, and exclude profanities for more readable translations. Deliver readable translations with an engine trained to normalize speech output.
Learn more
Translator Guru
Translator Guru is a mobile translation app designed to turn a smartphone into a real-time communication tool capable of translating speech, text, and images across more than 100 languages. It enables users to type, speak, or use the camera to translate content instantly, supporting scenarios like live conversations, reading menus or signs, and messaging across languages. It includes voice-to-voice and voice-to-speech conversation modes, allowing two people to communicate naturally in different languages with immediate playback of translated audio. It also integrates a translator keyboard that works across messaging apps, making it possible to translate text directly while chatting without switching tools. In addition to real-time translation, it offers built-in dictionaries and phrasebooks to help users understand meanings, pronunciation, and common expressions, along with features like saving favorites, viewing translation history, and sharing results.
Learn more