Mercury Edit 2Inception
|
Mercury VoiceInception
|
|||||
Related Products
|
||||||
About
Mercury Edit 2 is part of Inception Labs’ Mercury family of AI models, designed to perform high-speed reasoning, coding, and editing tasks using a fundamentally different architecture from traditional large language models. It builds on Mercury 2, a diffusion-based reasoning model that generates and refines entire outputs in parallel rather than producing text token by token, enabling significantly faster performance and more responsive editing workflows. Instead of acting like a sequential “typewriter,” the system behaves more like an editor, starting with a rough draft and iteratively improving it across multiple tokens at once, which allows for real-time interaction and rapid iteration in tasks such as code editing, content generation, and agent-based workflows. This architecture delivers throughput of up to around 1,000 tokens per second, making it several times faster than conventional models while maintaining competitive reasoning quality across benchmarks.
|
About
Mercury is a family of diffusion large language models built to deliver frontier LLM quality at significantly higher speeds, running at 1,000+ tokens per second on commercial NVIDIA GPUs for instant, in-the-flow AI applications. The models are OpenAI-compatible and designed as drop-in replacements for traditional LLMs, making them easier to integrate into existing AI stacks. Mercury 2.5 is the family’s most intelligent reasoning dLLM, built for complex applications where both quality and speed matter. It supports a 260K context window, reasoning, tool use, and structured output, with use cases including rapid coding iteration, agents and subagents, customer support, and enterprise search. Mercury Voice is optimized for voice agents and delivers time-to-first-token under 170 ms while supporting reasoning, tool use, structured output, and a 128K context window. It is suited to applications such as customer support, patient care, education, and gaming.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
|||||
Audience
Developers and AI teams who need ultra-fast, real-time AI for coding, editing, and agent workflows with low latency and high throughput
|
Audience
Developers and AI teams seeking to build fast, high-performance applications, agents, voice systems, and reasoning workflows with diffusion-based language models
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
|||||
API
Offers API
Supported
|
API
Offers API
Supported
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$0.25 per 1M input tokens
Free Version
Not Supported
Free Trial
Supported
|
Pricing
$0.04 per 1M tokens
Free Version
Supported
Free Trial
Not Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Supported
In Person
Not Supported
|
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Supported
In Person
Not Supported
|
|||||
Company InformationInception
United States
www.inceptionlabs.ai/blog/introducing-mercury-edit-2
|
Company InformationInception
United States
www.inceptionlabs.ai/models
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
Claude Haiku 4.5
Not Supported
Cline
Supported
Cursor
Supported
ElevenLabs
Supported
GPT-5.6 Luna
Not Supported
Gemini 3.5 Flash-Lite
Not Supported
Inception Labs
Supported
Kilo Code
Supported
LangChain
Supported
OpenAI
Not Supported
|
Integrations
Claude Haiku 4.5
Supported
Cline
Not Supported
Cursor
Not Supported
ElevenLabs
Not Supported
GPT-5.6 Luna
Supported
Gemini 3.5 Flash-Lite
Supported
Inception Labs
Not Supported
Kilo Code
Not Supported
LangChain
Not Supported
OpenAI
Supported
|
|||||
|
|
|