Mercury VoiceInception
|
Qwen3-MaxAlibaba
|
|||||
Related Products
|
||||||
About
Mercury is a family of diffusion large language models built to deliver frontier LLM quality at significantly higher speeds, running at 1,000+ tokens per second on commercial NVIDIA GPUs for instant, in-the-flow AI applications. The models are OpenAI-compatible and designed as drop-in replacements for traditional LLMs, making them easier to integrate into existing AI stacks. Mercury 2.5 is the family’s most intelligent reasoning dLLM, built for complex applications where both quality and speed matter. It supports a 260K context window, reasoning, tool use, and structured output, with use cases including rapid coding iteration, agents and subagents, customer support, and enterprise search. Mercury Voice is optimized for voice agents and delivers time-to-first-token under 170 ms while supporting reasoning, tool use, structured output, and a 128K context window. It is suited to applications such as customer support, patient care, education, and gaming.
|
About
Qwen3-Max is Alibaba’s latest trillion-parameter large language model, designed to push performance in agentic tasks, coding, reasoning, and long-context processing. It is built atop the Qwen3 family and benefits from the architectural, training, and inference advances introduced there; mixing thinker and non-thinker modes, a “thinking budget” mechanism, and support for dynamic mode switching based on complexity. The model reportedly processes extremely long inputs (hundreds of thousands of tokens), supports tool invocation, and exhibits strong performance on benchmarks in coding, multi-step reasoning, and agent benchmarks (e.g., Tau2-Bench). While its initial variant emphasizes instruction following (non-thinking mode), Alibaba plans to bring reasoning capabilities online to enable autonomous agent behavior. Qwen3-Max inherits multilingual support and extensive pretraining on trillions of tokens, and it is delivered via API interfaces compatible with OpenAI-style functions.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Supported
Mac
Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Supported
iPad
Supported
Android
Supported
Chromebook
Not Supported
|
|||||
Audience
Developers and AI teams seeking to build fast, high-performance applications, agents, voice systems, and reasoning workflows with diffusion-based language models
|
Audience
AI product teams and research groups building agentic systems seeking an AI model with improved performance in coding and agent capabilities
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
|||||
API
Offers API
Supported
|
API
Offers API
Supported
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$0.04 per 1M tokens
Free Version
Supported
Free Trial
Not Supported
|
Pricing
Free
Free Version
Supported
Free Trial
Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Supported
In Person
Not Supported
|
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
|||||
Company InformationInception
United States
www.inceptionlabs.ai/models
|
Company InformationAlibaba
Founded: 1999
China
qwen.ai
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
OpenAI
Supported
Alibaba Cloud
Not Supported
Claude Haiku 4.5
Supported
Constellation Gate AI
Not Supported
GPT-5.6 Luna
Supported
Gemini 3.5 Flash-Lite
Supported
OpenClaw
Not Supported
Qwen Studio
Not Supported
Shiori
Not Supported
Sup AI
Not Supported
|
Integrations
OpenAI
Supported
Alibaba Cloud
Supported
Claude Haiku 4.5
Not Supported
Constellation Gate AI
Supported
GPT-5.6 Luna
Not Supported
Gemini 3.5 Flash-Lite
Not Supported
OpenClaw
Supported
Qwen Studio
Supported
Shiori
Supported
Sup AI
Supported
|
|||||
|
|
|