Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Description

Mercury 2 represents a groundbreaking advancement in reasoning models, specifically designed for real-time voice interaction as it can quickly answer phone calls. Unlike traditional autoregressive models that leave callers in silence while generating responses one token at a time, Mercury 2 employs a diffusion large language model architecture capable of producing over 1000 tokens per second with standard NVIDIA GPUs. This remarkable speed allows it to complete a full reasoning process and begin speaking within a timeframe that aligns with natural conversational flow, effectively shortening the typical wait time from several seconds to approximately 300 milliseconds. The operational mechanism of Mercury models involves transforming clear text into noise, after which a conventional Transformer is trained to reverse this transformation and predict the original text across all positions at once. By utilizing a denoising approach that engages multiple tokens simultaneously, generation becomes more efficient, enabling speeds akin to custom silicon on NVIDIA H100s while improving responsiveness in voice applications. As a result, Mercury 2 not only enhances user experience but also sets a new standard for interactive voice technologies.

Description

Mercury 2.5 represents the pinnacle of production models from Inception, demonstrating a remarkable enhancement in quality compared to its predecessor, Mercury 2, all while upholding an impressive low-latency serving profile. It stands out as the most advanced diffusion language model currently available and is touted by Inception as the largest diffusion LLM ever developed. With a 40% boost in intelligence over Mercury 2, its performance aligns closely with that of cost-efficient frontier models like GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5. The model boasts a generation speed of 1,107 tokens per second on commonly accessible NVIDIA GPUs and accommodates a generous 260K-token context window. Among its features are adjustable reasoning capabilities, simultaneous tool calls, and JSON that aligns with schemas. Specifically engineered for latency-sensitive tasks, it is well-suited for scenarios involving numerous model calls during a single interaction. In applications such as search agents and RAG pipelines, Mercury 2.5 excels in functions like planning, query rewriting, re-ranking, fact structuring, source summarization, and answer verification, all while ensuring rapid response times, making it an essential tool for developers seeking efficiency in their workflows.

API Access

Has API

API Access

Has API

Screenshots View All

Screenshots View All

Integrations

Cerebras
GPT-4.1
Groq
Inception Labs
JSON
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI

Integrations

Cerebras
GPT-4.1
Groq
Inception Labs
JSON
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI

Pricing Details

No price information available.
Free Trial
Free Version

Pricing Details

No price information available.
Free Trial
Free Version

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Vendor Details

Company Name

Inception

Country

United States

Website

www.inceptionlabs.ai/blog/mercury-2-the-first-reasoning-model-fast-enough-to-pick-up-the-phone

Vendor Details

Company Name

Inception

Country

United States

Website

www.inceptionlabs.ai/blog/introducing-mercury-2-5

Product Features

Product Features

Alternatives

Mercury Coder Reviews

Mercury Coder

Inception Labs

Alternatives

Mercury Edit 2 Reviews

Mercury Edit 2

Inception
Mercury Coder Reviews

Mercury Coder

Inception Labs
Mercury 2.5 Reviews

Mercury 2.5

Inception
Mercury Edit 2 Reviews

Mercury Edit 2

Inception
Mercury 2 Reviews

Mercury 2

Inception