Mercury 2.5Inception
|
||||||
Related Products
|
||||||
About
Mercury 2.5 is Inception’s most capable production model yet and a significant step up in quality over Mercury 2 while maintaining the same low-latency serving profile. It is the most capable diffusion LLM on the market and, according to Inception, the largest diffusion language model ever trained. Mercury 2.5 delivers a 40% increase in intelligence over Mercury 2, with performance comparable to cost-optimized frontier models such as GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5. It generates at 1,107 tokens per second on widely available NVIDIA GPUs and supports a 260K-token context window. Capabilities include tunable reasoning, parallel tool calls, and schema-aligned JSON. The model is designed for latency-sensitive workloads where many model calls may happen inside a single interaction. In search agents and RAG pipelines, it can support planning, query rewriting, reranking, fact structuring, source summarization, and answer checking while keeping calls fast.
|
About
Effective self-service and faster responses to your customers' questions, around the clock, lead to measurable improvements in your customer satisfaction. Dramatically increase your conversion rate with personalized product advice, suggestions, and purchase decision simplification. Automate your service effectively and have requests resolved before they become tickets. This significantly reduces the workload of your service team. Mercury's unique dialog technology enables an unmatched level of context, personalization, and intelligent dialog. This pays off for you by enabling more complex use cases and noticeably better UX. The active learning behavior combines two decisive advantages: It increases dialog robustness and guides your customers to the goal even with difficult questions, while leading to independent improvement of language comprehension.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Developers and enterprises seeking a high-speed diffusion language model for search, RAG, voice agents, coding assistants, and latency-sensitive AI applications
|
Audience
Ecommerce and service teams wanting a solution providing AI chatbots and messaging tools
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
No information available.
Free Version
Free Trial
|
Pricing
€500 per month
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationInception
United States
www.inceptionlabs.ai/blog/introducing-mercury-2-5
|
Company InformationMercury
Founded: 2016
Germany
www.mercury.ai/
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
||||||
Categories |
Categories |
|||||
|
|
|