JevTypeSafe AI
|
Mercury 2Inception
|
|||||
Related Products
|
||||||
About
Jev is a System One Model from TypeSafe AI designed for fast, structured decision-making inside software and automated workflows. Unlike conventional large language models that generate free-form text token by token, Jev accepts unstructured state and returns predefined type-safe values with calibrated probabilities and confidence scores. The model uses a parallel sampling architecture and a training approach called Reinforcement Learning for Calibrated Decisions, which is optimized for producing reliable probabilistic decisions rather than conversational responses. Jev is intended for tasks such as classification, routing, scoring, extraction, branching logic, verification, guardrails, and other AI-powered workflows where structured outputs can plug directly into application code. TypeSafe reports end-to-end response times ranging from about 70 to 500 milliseconds and says Jev can be substantially faster and more efficient than frontier language models on System One-style workloads.
|
About
Mercury 2 is the first reasoning model fast enough to pick up the phone, a reasoning diffusion language model built for real-time voice agents. Instead of making callers wait through seconds of dead air while an autoregressive model generates thinking tokens one by one, Mercury 2 uses a diffusion large language model architecture to generate tokens in parallel, decoding 1000+ tokens per second on standard NVIDIA GPUs. That speed is fast enough to run a full reasoning pass and start speaking within the latency budget of a natural conversation, reducing the cost of reasoning from seconds of silence to roughly 300 milliseconds. Mercury models work by corrupting clean text into noise, then training a standard Transformer to reverse the process and predict clean text across all positions simultaneously. Because each denoising pass touches many tokens, generation uses the GPU more efficiently than one-token-at-a-time decoding, making custom-silicon-like speed possible on NVIDIA H100s.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Software developers, AI engineers, automation teams, infrastructure builders, and enterprises that need low-latency, type-safe, probabilistic AI decisions for production software and structured workflows
|
Audience
Voice AI developers and product teams seeking a fast reasoning model for low-latency conversational agents that still handle complex instructions
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
Input: $0.042 / 1M tokens
Input tokens: $0.042 / 1 million tokens ($42 per billion tokens).
Output tokens: FREE (too cheap to meter).
Free Version
Free Trial
|
Pricing
No information available.
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationTypeSafe AI
Founded: 2024
United States
typesafe.ai/
|
Company InformationInception
United States
www.inceptionlabs.ai/blog/mercury-2-the-first-reasoning-model-fast-enough-to-pick-up-the-phone
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
||||||
|
|
||||||
Categories |
Categories |
|||||
Integrations
Cerebras
GPT-4.1
Groq
Inception Labs
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI
|
Integrations
Cerebras
GPT-4.1
Groq
Inception Labs
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI
|
|||||
|
|
|