Gemini 3.8 LiveGoogle
|
Gemini Live APIGoogle
|
|||||
Related Products
|
||||||
About
Gemini 3.8 Live is Google DeepMind’s real-time speech-to-speech AI model for building conversational voice applications and interactive agents. The model can maintain natural dialogue while reasoning, using tools, and carrying out tasks during an ongoing conversation. Asynchronous function calling allows applications to execute API and tool requests in the background while Gemini continues streaming audio responses to the user. Gemini 3.8 Live can also incorporate live visual context, enabling agents to respond based on what users say and what the system can see. It supports more than 97 languages, maintains accent consistency, and is designed to accurately interpret alphanumeric information such as confirmation codes, claim numbers, and technical data. Gemini 3.8 Live is available through the Gemini Live API and Google AI Studio for developers building customer service agents, assistants, training applications, and other voice-first experiences.
|
About
The Gemini Live API is a preview feature that enables low-latency, bidirectional voice and video interactions with Gemini. It allows end users to experience natural, human-like voice conversations and provides the ability to interrupt the model's responses using voice commands. The model can process text, audio, and video input, and it can provide text and audio output. New capabilities include two new voices and 30 new languages with configurable output language, configurable image resolutions (66/256 tokens), configurable turn coverage (send all inputs all the time or only when the user is speaking), configurable interruption settings, configurable voice activity detection, new client events for end-of-turn signaling, token counts, a client event for signaling the end of stream, text streaming, configurable session resumption with session data stored on the server for 24 hours, and longer session support with a sliding context window.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Developers, AI application teams, enterprises, contact centers, voice-agent builders, and software companies creating real-time conversational applications that need multilingual speech, visual understanding, reasoning, and tool use
|
Audience
Researchers looking for a solution to build real-time, multimodal AI applications that require low-latency voice and video interactions
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
No information available.
Free Version
Free Trial
|
Pricing
No information available.
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationGoogle
Founded: 1998
United States
gemini.google.com
|
Company InformationGoogle
Founded: 1998
United States
ai.google.dev/gemini-api/docs/live
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
||||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
Agora
Fishjam
Gemini
Gemini Enterprise
Gemini Enterprise Agent Platform
Google AI Studio
Google Stitch
LiveKit
Vision Agents
Firebase
|
Integrations
Agora
Fishjam
Gemini
Gemini Enterprise
Gemini Enterprise Agent Platform
Google AI Studio
Google Stitch
LiveKit
Vision Agents
Firebase
|
|||||
|
|
|