Gemini DiffusionGoogle DeepMind
|
Mercury 2.5Inception
|
|||||
Related Products
|
||||||
About
Gemini Diffusion is our state-of-the-art research model exploring what diffusion means for language and text generation. Large-language models are the foundation of generative AI today. We’re using a technique called diffusion to explore a new kind of language model that gives users greater control, creativity, and speed in text generation. Diffusion models work differently. Instead of predicting text directly, they learn to generate outputs by refining noise, step by step. This means they can iterate on a solution very quickly and error correct during the generation process. This helps them excel at tasks like editing, including in the context of math and code. Generates entire blocks of tokens at once, meaning it responds more coherently to a user’s prompt than autoregressive models. Gemini Diffusion’s external benchmark performance is comparable to much larger models, whilst also being faster.
|
About
Mercury 2.5 is Inception’s most capable production model yet and a significant step up in quality over Mercury 2 while maintaining the same low-latency serving profile. It is the most capable diffusion LLM on the market and, according to Inception, the largest diffusion language model ever trained. Mercury 2.5 delivers a 40% increase in intelligence over Mercury 2, with performance comparable to cost-optimized frontier models such as GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5. It generates at 1,107 tokens per second on widely available NVIDIA GPUs and supports a 260K-token context window. Capabilities include tunable reasoning, parallel tool calls, and schema-aligned JSON. The model is designed for latency-sensitive workloads where many model calls may happen inside a single interaction. In search agents and RAG pipelines, it can support planning, query rewriting, reranking, fact structuring, source summarization, and answer checking while keeping calls fast.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
AI researchers and developers seeking a tool providing editable text generation by leveraging diffusion-based language modeling
|
Audience
Developers and enterprises seeking a high-speed diffusion language model for search, RAG, voice agents, coding assistants, and latency-sensitive AI applications
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
No information available.
Free Version
Free Trial
|
Pricing
No information available.
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationGoogle DeepMind
Founded: 2010
United Kingdom
deepmind.google/models/gemini-diffusion/
|
Company InformationInception
United States
www.inceptionlabs.ai/blog/introducing-mercury-2-5
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
Categories |
Categories |
|||||
|
|
|