Qwen3-TTS

Qwen3-TTS

Alibaba
+
+

Related Products

  • Google Cloud Speech-to-Text
    366 Ratings
    Visit Website
  • QEval
    30 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Squaretalk
    293 Ratings
    Visit Website
  • Assembled
    272 Ratings
    Visit Website
  • Forethought
    166 Ratings
    Visit Website
  • Dialpad Support
    1,600 Ratings
    Visit Website
  • Phonexa
    238 Ratings
    Visit Website
  • DialerAI
    5 Ratings
    Visit Website
  • LALAL.AI
    5,230 Ratings
    Visit Website

About

Boson AI provides voice agents powered by foundation audio models, built to run in business workflows and learn from every call. Higgs Realtime enables live voice agents for support lines, sales calls, and product assistants that listen, reason, call tools, and respond in real time with low latency and natural speech-to-speech interaction. Higgs Audio and Avatar extend these capabilities with text-to-speech, speech-to-text, voice cloning, sentiment detection, and avatar generation, producing natural speech while understanding tone, emotion, and intent. The models support high-accuracy multilingual speech recognition, real-time translation, and expressive voice generation, while sentiment signals can improve routing, analytics, and context-aware agent behavior. Designed for real-world production, the platform emphasizes quality, latency, reliability, and flexible deployment across managed and self-serve environments.

About

Qwen3-TTS is an open source series of advanced text-to-speech models developed by the Qwen team at Alibaba Cloud under the Apache-2.0 license, offering stable, expressive, and real-time speech generation with features such as voice cloning, voice design, and fine-grained control of prosody and acoustic attributes. The models support 10 major languages, including Chinese, English, Japanese, Korean, German, French, Russian, Portuguese, Spanish, and Italian, and multiple dialectal voice profiles with adaptive control over tone, speaking rate, and emotional expression based on text semantics and instructions. Qwen3-TTS uses efficient tokenization and a dual-track architecture that enables ultra-low-latency streaming synthesis (first audio packet in ~97 ms), making it suitable for interactive and real-time use cases, and includes a range of models with different capabilities (e.g., rapid 3-second voice cloning, custom voice timbres, and instruction-based voice design).

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Enterprises, developers, and product teams that want to build and deploy natural, real-time voice agents that automate conversations and complete business workflows

Audience

Researchers who need a model for expressive, multilingual, controllable, and streaming voice generation in applications like voice assistants, dubbing, accessibility, and creative audio synthesis

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Boson AI
Founded: 2023
United States
www.boson.ai/

Company Information

Alibaba
Founded: 1999
China
github.com/QwenLM/Qwen3-TTS

Alternatives

Alternatives

Simba 3.2

Simba 3.2

Speechify
CosyVoice

CosyVoice

Alibaba
Higgs Realtime

Higgs Realtime

Boson AI
MAI-Voice-2

MAI-Voice-2

Microsoft AI

Categories

Categories

Integrations

Alibaba Cloud
Higgs Audio / Avatar
Higgs Realtime
OpenClaw
Qwen

Integrations

Alibaba Cloud
Higgs Audio / Avatar
Higgs Realtime
OpenClaw
Qwen
Claim Boson AI and update features and information
Claim Boson AI and update features and information
Claim Qwen3-TTS and update features and information
Claim Qwen3-TTS and update features and information