Mercury 2.5

Mercury 2.5

Inception
+
+

Related Products

  • LTX
    182 Ratings
    Visit Website
  • Runpod
    230 Ratings
    Visit Website
  • FinOpsly
    3 Ratings
    Visit Website
  • Dragonfly
    16 Ratings
    Visit Website
  • AlsoThere
    1 Rating
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Chainstack
    34 Ratings
    Visit Website
  • Innoslate
    93 Ratings
    Visit Website
  • CallHub
    427 Ratings
    Visit Website
  • Teradata VantageCloud
    1,121 Ratings
    Visit Website

About

Mercury 2.5 is Inception’s most capable production model yet and a significant step up in quality over Mercury 2 while maintaining the same low-latency serving profile. It is the most capable diffusion LLM on the market and, according to Inception, the largest diffusion language model ever trained. Mercury 2.5 delivers a 40% increase in intelligence over Mercury 2, with performance comparable to cost-optimized frontier models such as GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5. It generates at 1,107 tokens per second on widely available NVIDIA GPUs and supports a 260K-token context window. Capabilities include tunable reasoning, parallel tool calls, and schema-aligned JSON. The model is designed for latency-sensitive workloads where many model calls may happen inside a single interaction. In search agents and RAG pipelines, it can support planning, query rewriting, reranking, fact structuring, source summarization, and answer checking while keeping calls fast.

About

​Reka Flash 3 is a 21-billion-parameter multimodal AI model developed by Reka AI, designed to excel in general chat, coding, instruction following, and function calling. It processes and reasons with text, images, video, and audio inputs, offering a compact, general-purpose solution for various applications. Trained from scratch on diverse datasets, including publicly accessible and synthetic data, Reka Flash 3 underwent instruction tuning on curated, high-quality data to optimize performance. The final training stage involved reinforcement learning using REINFORCE Leave One-Out (RLOO) with both model-based and rule-based rewards, enhancing its reasoning capabilities. With a context length of 32,000 tokens, Reka Flash 3 performs competitively with proprietary models like OpenAI's o1-mini, making it suitable for low-latency or on-device deployments. The model's full precision requires 39GB (fp16), but it can be compressed to as small as 11GB using 4-bit quantization.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Developers and enterprises seeking a high-speed diffusion language model for search, RAG, voice agents, coding assistants, and latency-sensitive AI applications

Audience

Developers seeking an AI model for coding assistance, natural language understanding, and multimodal data processing in their applications

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Inception
United States
www.inceptionlabs.ai/blog/introducing-mercury-2-5

Company Information

Reka
Founded: 2022
United States
www.reka.ai/news/introducing-reka-flash

Alternatives

Mercury Coder

Mercury Coder

Inception Labs

Alternatives

Smaug Flash

Smaug Flash

Abacus.AI
Mercury Edit 2

Mercury Edit 2

Inception
Mercury 2

Mercury 2

Inception
Gemini 4

Gemini 4

Google
GLM-4.1V

GLM-4.1V

Z.ai

Categories

Categories

Integrations

JSON
Nexus
Space

Integrations

JSON
Nexus
Space
Claim Mercury 2.5 and update features and information
Claim Mercury 2.5 and update features and information
Claim Reka Flash 3 and update features and information
Claim Reka Flash 3 and update features and information