Mercury VoiceInception
|
MiniMax M1MiniMax
|
|||||
Related Products
|
||||||
About
Mercury is a family of diffusion large language models built to deliver frontier LLM quality at significantly higher speeds, running at 1,000+ tokens per second on commercial NVIDIA GPUs for instant, in-the-flow AI applications. The models are OpenAI-compatible and designed as drop-in replacements for traditional LLMs, making them easier to integrate into existing AI stacks. Mercury 2.5 is the family’s most intelligent reasoning dLLM, built for complex applications where both quality and speed matter. It supports a 260K context window, reasoning, tool use, and structured output, with use cases including rapid coding iteration, agents and subagents, customer support, and enterprise search. Mercury Voice is optimized for voice agents and delivers time-to-first-token under 170 ms while supporting reasoning, tool use, structured output, and a 128K context window. It is suited to applications such as customer support, patient care, education, and gaming.
|
About
MiniMax‑M1 is a large‑scale hybrid‑attention reasoning model released by MiniMax AI under the Apache 2.0 license. It supports an unprecedented 1 million‑token context window and up to 80,000-token outputs, enabling extended reasoning across long documents. Trained using large‑scale reinforcement learning with a novel CISPO algorithm, MiniMax‑M1 completed full training on 512 H800 GPUs in about three weeks. It achieves state‑of‑the‑art performance on benchmarks in mathematics, coding, software engineering, tool usage, and long‑context understanding, matching or outperforming leading models. Two model variants are available (40K and 80K thinking budgets), with weights and deployment scripts provided via GitHub and Hugging Face.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
|||||
Audience
Developers and AI teams seeking to build fast, high-performance applications, agents, voice systems, and reasoning workflows with diffusion-based language models
|
Audience
AI researchers, developers, and enterprises needing a solution providing LLM capable of long‑context reasoning, efficient compute, and integration via function calls
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
|||||
API
Offers API
Supported
|
API
Offers API
Supported
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$0.04 per 1M tokens
Free Version
Supported
Free Trial
Not Supported
|
Pricing
No information available.
Free Version
Not Supported
Free Trial
Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Supported
In Person
Not Supported
|
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
|||||
Company InformationInception
United States
www.inceptionlabs.ai/models
|
Company InformationMiniMax
Founded: 2021
Singapore
www.minimax.io/news/minimaxm1
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
Anuma
Not Supported
Claude Haiku 4.5
Supported
GPT-5.6 Luna
Supported
Gemini 3.5 Flash-Lite
Supported
GitHub
Not Supported
Hugging Face
Not Supported
OpenAI
Supported
SiliconFlow
Not Supported
|
Integrations
Anuma
Supported
Claude Haiku 4.5
Not Supported
GPT-5.6 Luna
Not Supported
Gemini 3.5 Flash-Lite
Not Supported
GitHub
Supported
Hugging Face
Supported
OpenAI
Not Supported
SiliconFlow
Supported
|
|||||
|
|
|