EVI 3

EVI 3

Hume AI
+
+

Related Products

  • Google Cloud Speech-to-Text
    366 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Google AI Studio
    41 Ratings
    Visit Website
  • QEval
    30 Ratings
    Visit Website
  • Adobe Firefly
    25,030 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    999 Ratings
    Visit Website
  • LTX
    182 Ratings
    Visit Website
  • IONOS Cloud GPU Servers
    45,199 Ratings
    Visit Website
  • JetBrains Junie
    12 Ratings
    Visit Website
  • LALAL.AI
    5,355 Ratings
    Visit Website

About

Hume AI's EVI 3 is a third-generation speech-language model that streams in user speech and forms natural, expressive speech and language responses. At conversational latency, it produces the same quality of speech as our text-to-speech model, Octave. Simultaneously, it responds with the same intelligence as the most advanced LLMs of similar latency. It also communicates with reasoning models and web search systems as it speaks, “thinking fast and slow” to match the intelligence of any frontier AI system. EVI 3 can instantly generate new voices and personalities instead of being limited to a handful of speakers. For instance, users can speak to any of the more than 100,000 custom voices already created on our text-to-speech platform, each with an inferred personality. No matter the voice, it responds with a wide range of emotions or styles, implicitly or on command.

About

MAI-Transcribe-2-Streaming is a low-latency streaming transcription model built for real-time speech applications, delivering transcripts in 60 languages with automatic, continuous language detection. Rather than waiting for someone to finish speaking, it produces initial partial transcripts just over 100 ms after receiving audio, revises them as more context arrives, and commits stable text quickly. This allows voice applications to begin reasoning, calling tools, or displaying live transcripts while a person is still speaking. Microsoft reports that the model ranks No. 1 for both final and partial transcript accuracy on Artificial Analysis. MAI-Voice-2.1 complements it with multilingual text-to-speech across 23 languages and 26 locales, allowing a single voice to switch languages while maintaining the same speaker identity and adopting native accents.

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Not Supported
iPhone Supported
iPad Supported
Android Not Supported
Chromebook Not Supported

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Not Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Audience

Developers and businesses in search of a solution to integrate emotionally intelligent, real-time voice AI into their applications

Audience

Developers and AI teams seeking to build fast, multilingual voice agents, real-time transcription systems, and conversational speech applications

Support

Phone Support Not Supported
24/7 Live Support Not Supported
Online Supported

Support

Phone Support Not Supported
24/7 Live Support Not Supported
Online Supported

API

Offers API Supported

API

Offers API Not Supported

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version Supported
Free Trial Not Supported

Pricing

No information available.
Free Version Not Supported
Free Trial Not Supported

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation Supported
Webinars Not Supported
Live Online Not Supported
In Person Not Supported

Training

Documentation Supported
Webinars Not Supported
Live Online Not Supported
In Person Not Supported

Company Information

Hume AI
Founded: 2021
United States
www.hume.ai/blog/introducing-evi-3

Company Information

Microsoft AI
Founded: 2024
United States
microsoft.ai/news/our-first-streaming-transcription-model/

Alternatives

Octave TTS

Octave TTS

Hume AI

Alternatives

Azure AI Speech

Azure AI Speech

Microsoft
Cartesia Ink 2

Cartesia Ink 2

Cartesia

Categories

Categories

AI Models Supported
Speech to Text Supported

Integrations

Hume AI Supported

Integrations

Hume AI Not Supported
Claim EVI 3 and update features and information
Claim EVI 3 and update features and information
Claim MAI-Transcribe-2-Streaming and update features and information
Claim MAI-Transcribe-2-Streaming and update features and information