Mercury 2

Mercury 2

Inception
Olmo 3

Olmo 3

Ai2
+
+

Related Products

  • Checksum.ai
    1 Rating
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • ND Wallet
    14 Ratings
    Visit Website
  • Runpod
    220 Ratings
    Visit Website
  • Google AI Studio
    30 Ratings
    Visit Website
  • Imorgon
    5 Ratings
    Visit Website
  • Qloo
    23 Ratings
    Visit Website
  • MEXC
    188,765 Ratings
    Visit Website
  • EBizCharge
    207 Ratings
    Visit Website
  • IUX
    927 Ratings
    Visit Website

About

Mercury 2 is the first reasoning model fast enough to pick up the phone, a reasoning diffusion language model built for real-time voice agents. Instead of making callers wait through seconds of dead air while an autoregressive model generates thinking tokens one by one, Mercury 2 uses a diffusion large language model architecture to generate tokens in parallel, decoding 1000+ tokens per second on standard NVIDIA GPUs. That speed is fast enough to run a full reasoning pass and start speaking within the latency budget of a natural conversation, reducing the cost of reasoning from seconds of silence to roughly 300 milliseconds. Mercury models work by corrupting clean text into noise, then training a standard Transformer to reverse the process and predict clean text across all positions simultaneously. Because each denoising pass touches many tokens, generation uses the GPU more efficiently than one-token-at-a-time decoding, making custom-silicon-like speed possible on NVIDIA H100s.

About

Olmo 3 is a fully open model family spanning 7 billion and 32 billion parameter variants that delivers not only high-performing base, reasoning, instruction, and reinforcement-learning models, but also exposure of the entire model flow, including raw training data, intermediate checkpoints, training code, long-context support (65,536 token window), and provenance tooling. Starting with the Dolma 3 dataset (≈9 trillion tokens) and its disciplined mix of web text, scientific PDFs, code, and long-form documents, the pre-training, mid-training, and long-context phases shape the base models, which are then post-trained via supervised fine-tuning, direct preference optimisation, and RL with verifiable rewards to yield the Think and Instruct variants. The 32 B Think model is described as the strongest fully open reasoning model to date, competitively close to closed-weight peers in math, code, and complex reasoning.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Voice AI infrastructure teams that need low-latency reasoning models for phone agents, tool-calling workflows, and natural customer conversations

Audience

AI researchers, developers and enterprises needing a tool offering foundation models to inspect, fine-tune or deploy with full provenance and auditability

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Inception
United States
www.inceptionlabs.ai/blog/mercury-2-the-first-reasoning-model-fast-enough-to-pick-up-the-phone

Company Information

Ai2
Founded: 2014
United States
allenai.org/blog/olmo3

Alternatives

Mercury Coder

Mercury Coder

Inception Labs

Alternatives

LongCat-2.0

LongCat-2.0

LongCat
Mercury Edit 2

Mercury Edit 2

Inception
Qwen3-Max

Qwen3-Max

Alibaba
ByteDance Seed

ByteDance Seed

ByteDance
MiniMax M1

MiniMax M1

MiniMax
Uni-1

Uni-1

Luma AI
DeepSeek-V4

DeepSeek-V4

DeepSeek

Categories

Categories

Integrations

Cerebras
GPT-4.1
Groq
Inception Labs
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI

Integrations

Cerebras
GPT-4.1
Groq
Inception Labs
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI
Claim Mercury 2 and update features and information
Claim Mercury 2 and update features and information
Claim Olmo 3 and update features and information
Claim Olmo 3 and update features and information