Ratings and Reviews 0 Ratings

Total
ease
features
design
support

This software has no reviews. Be the first to write a review.

Write a Review

Ratings and Reviews 0 Ratings

Total
ease
features
design
support

This software has no reviews. Be the first to write a review.

Write a Review

Alternatives to Consider

  • LTX Reviews & Ratings
    182 Ratings
    Company Website
  • Runpod Reviews & Ratings
    230 Ratings
    Company Website
  • Dragonfly Reviews & Ratings
    16 Ratings
    Company Website
  • AlsoThere Reviews & Ratings
    1 Rating
    Company Website
  • LM-Kit.NET Reviews & Ratings
    29 Ratings
    Company Website
  • Chainstack Reviews & Ratings
    34 Ratings
    Company Website
  • Innoslate Reviews & Ratings
    93 Ratings
    Company Website
  • CallTools Reviews & Ratings
    641 Ratings
    Company Website
  • CallHub Reviews & Ratings
    427 Ratings
    Company Website
  • Teradata VantageCloud Reviews & Ratings
    1,121 Ratings
    Company Website

What is Mercury 2.5?

Mercury 2.5 exemplifies the ultimate achievement in production models from Inception, showcasing a significant quality improvement over its predecessor, Mercury 2, while maintaining an outstanding low-latency performance. This model is recognized as the most sophisticated diffusion language model currently on the market and is claimed by Inception to be the largest diffusion LLM ever created. With a remarkable 40% increase in intelligence compared to Mercury 2, its capabilities closely mirror those of economical frontier models such as GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5. It features an impressive generation rate of 1,107 tokens per second on widely used NVIDIA GPUs and can handle a substantial 260K-token context window. Among its many attributes are customizable reasoning abilities, concurrent tool executions, and JSON compatibility with schemas. Designed specifically for tasks sensitive to latency, it excels in environments that require multiple model calls during one interaction. In various applications, including search agents and RAG pipelines, Mercury 2.5 performs exceptionally well in activities such as planning, query rewriting, re-ranking, fact structuring, source summarization, and answer verification. This efficiency ensures swift response times, making it a vital asset for developers aiming to enhance their workflow productivity. As technology continues to evolve, models like Mercury 2.5 will likely set new benchmarks in the industry.

What is Mercury 2?

Mercury 2 signifies a revolutionary leap in reasoning models, particularly tailored for instantaneous voice interactions, as it can promptly respond to incoming calls. In contrast to conventional autoregressive models that often leave callers waiting in silence while they generate responses sequentially, Mercury 2 uses a diffusion large language model architecture that can produce more than 1000 tokens per second on standard NVIDIA GPUs. This extraordinary processing speed enables it to finalize a complete reasoning cycle and start speaking in a timeframe that harmonizes with the natural flow of conversation, effectively reducing the usual wait time from several seconds to around 300 milliseconds. The functionality of Mercury models revolves around converting clear text into noise, after which a traditional Transformer is trained to reverse this process and predict the original text simultaneously across all positions. By adopting a denoising strategy that processes multiple tokens concurrently, the generation process becomes more efficient, achieving speeds comparable to customized silicon on NVIDIA H100s while enhancing responsiveness in voice applications. Consequently, Mercury 2 not only improves user interactions but also establishes a new benchmark for the field of interactive voice technology, paving the way for future advancements. With its innovative design, it promises to revolutionize the way users engage with voice systems.

Media

Media

Integrations Supported

Cerebras
GPT-4.1
Groq
Inception Labs
JSON
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI

Integrations Supported

Cerebras
GPT-4.1
Groq
Inception Labs
JSON
LiveKit
OpenAI
Pipecat
Retell AI
Vapi AI

API Availability

Has API

API Availability

Has API

Pricing Information

Pricing not provided
Free Version
Free Trial Offered?

Pricing Information

Pricing not provided
Free Version
Free Trial Offered?

Supported Platforms

SaaS
Android
iPhone
iPad
Windows
Mac
On-Prem
Chromebook
Linux

Supported Platforms

SaaS
Android
iPhone
iPad
Windows
Mac
On-Prem
Chromebook
Linux

Customer Service / Support

Standard Support
24 Hour Support
Web-Based Support

Customer Service / Support

Standard Support
24 Hour Support
Web-Based Support

Training Options

Documentation Hub
Webinars
Online Training
On-Site Training

Training Options

Documentation Hub
Webinars
Online Training
On-Site Training

Company Facts

Organization Name

Inception

Company Location

United States

Company Website

www.inceptionlabs.ai/blog/introducing-mercury-2-5

Company Facts

Organization Name

Inception

Company Location

United States

Company Website

www.inceptionlabs.ai/blog/mercury-2-the-first-reasoning-model-fast-enough-to-pick-up-the-phone

Categories and Features

Categories and Features

Popular Alternatives

Mercury Coder Reviews & Ratings

Mercury Coder

Inception Labs

Popular Alternatives

Mercury Edit 2 Reviews & Ratings

Mercury Edit 2

Inception
Mercury Coder Reviews & Ratings

Mercury Coder

Inception Labs
Mercury 2 Reviews & Ratings

Mercury 2

Inception
Mercury Edit 2 Reviews & Ratings

Mercury Edit 2

Inception
Mercury 2.5 Reviews & Ratings

Mercury 2.5

Inception