ByteDance Seed

ByteDance Seed

ByteDance
Mercury 2

Mercury 2

Inception
+
+

Related Products

  • Checksum.ai
    1 Rating
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Innoslate
    93 Ratings
    Visit Website
  • Imorgon
    5 Ratings
    Visit Website
  • Runpod
    220 Ratings
    Visit Website
  • Epicor Kinetic
    533 Ratings
    Visit Website
  • KonstructIQ
    7 Ratings
    Visit Website
  • Knak
    166 Ratings
    Visit Website
  • EBizCharge
    207 Ratings
    Visit Website
  • Apryse PDF SDK
    157 Ratings
    Visit Website

About

Seed Diffusion Preview is a large-scale, code-focused language model that uses discrete-state diffusion to generate code non-sequentially, achieving dramatically faster inference without sacrificing quality by decoupling generation from the token-by-token bottleneck of autoregressive models. It combines a two-stage curriculum, mask-based corruption followed by edit-based augmentation, to robustly train a standard dense Transformer, striking a balance between speed and accuracy and avoiding shortcuts like carry-over unmasking to preserve principled density estimation. The model delivers an inference speed of 2,146 tokens/sec on H20 GPUs, outperforming contemporary diffusion baselines while matching or exceeding their accuracy on standard code benchmarks, including editing tasks, thereby establishing a new speed-quality Pareto frontier and demonstrating discrete diffusion’s practical viability for real-world code generation.

About

Mercury 2 is the first reasoning model fast enough to pick up the phone, a reasoning diffusion language model built for real-time voice agents. Instead of making callers wait through seconds of dead air while an autoregressive model generates thinking tokens one by one, Mercury 2 uses a diffusion large language model architecture to generate tokens in parallel, decoding 1000+ tokens per second on standard NVIDIA GPUs. That speed is fast enough to run a full reasoning pass and start speaking within the latency budget of a natural conversation, reducing the cost of reasoning from seconds of silence to roughly 300 milliseconds. Mercury models work by corrupting clean text into noise, then training a standard Transformer to reverse the process and predict clean text across all positions simultaneously. Because each denoising pass touches many tokens, generation uses the GPU more efficiently than one-token-at-a-time decoding, making custom-silicon-like speed possible on NVIDIA H100s.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Developers and research teams needing a code generation model tool that reduces latency compared to traditional autoregressive systems while maintaining competitive benchmark performance

Audience

Voice AI infrastructure teams that need low-latency reasoning models for phone agents, tool-calling workflows, and natural customer conversations

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

ByteDance
Founded: 2012
China
seed.bytedance.com/en/seed_diffusion

Company Information

Inception
United States
www.inceptionlabs.ai/blog/mercury-2-the-first-reasoning-model-fast-enough-to-pick-up-the-phone

Alternatives

Seed2.0 Pro

Seed2.0 Pro

ByteDance

Alternatives

Mercury Coder

Mercury Coder

Inception Labs
Mercury Edit 2

Mercury Edit 2

Inception
Gemini Diffusion

Gemini Diffusion

Google DeepMind
Mercury Coder

Mercury Coder

Inception Labs
ByteDance Seed

ByteDance Seed

ByteDance
Mercury 2

Mercury 2

Inception
Uni-1

Uni-1

Luma AI

Categories

Categories

Integrations

AnyAPI
C++
Citlyze
Flyne AI
Fuser
GPT-4.1
Groq
Inception Labs
Java
LiveKit
Magica
OpenAI
Pipecat
Python
Retell AI
TESS AI
Vapi AI
WaveSpeedAI
ZOOOP
graphis

Integrations

AnyAPI
C++
Citlyze
Flyne AI
Fuser
GPT-4.1
Groq
Inception Labs
Java
LiveKit
Magica
OpenAI
Pipecat
Python
Retell AI
TESS AI
Vapi AI
WaveSpeedAI
ZOOOP
graphis
Claim ByteDance Seed and update features and information
Claim ByteDance Seed and update features and information
Claim Mercury 2 and update features and information
Claim Mercury 2 and update features and information