Mercury 2.5 Description

Mercury 2.5 represents the pinnacle of production models from Inception, demonstrating a remarkable enhancement in quality compared to its predecessor, Mercury 2, all while upholding an impressive low-latency serving profile. It stands out as the most advanced diffusion language model currently available and is touted by Inception as the largest diffusion LLM ever developed. With a 40% boost in intelligence over Mercury 2, its performance aligns closely with that of cost-efficient frontier models like GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5. The model boasts a generation speed of 1,107 tokens per second on commonly accessible NVIDIA GPUs and accommodates a generous 260K-token context window. Among its features are adjustable reasoning capabilities, simultaneous tool calls, and JSON that aligns with schemas. Specifically engineered for latency-sensitive tasks, it is well-suited for scenarios involving numerous model calls during a single interaction. In applications such as search agents and RAG pipelines, Mercury 2.5 excels in functions like planning, query rewriting, re-ranking, fact structuring, source summarization, and answer verification, all while ensuring rapid response times, making it an essential tool for developers seeking efficiency in their workflows.

Integrations

API:
Yes, Mercury 2.5 has an API

Reviews

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Company Details

Company:
Inception
Headquarters:
United States
Website:
www.inceptionlabs.ai/blog/introducing-mercury-2-5

Media

Mercury 2.5 Screenshot 1
Recommended Products
PRTG Catches Network Issues Before They Cause Downtime Icon
PRTG Catches Network Issues Before They Cause Downtime

Threshold-based alerts flag problems early, so your team can act before users notice, not after.

Reactive troubleshooting usually means hearing about a problem from frustrated users, not your monitoring tool. PRTG sets threshold-based alerts across devices, servers and applications, notifying your team by email, SMS or push the moment a metric crosses a set limit. That means catching a failing disk or overloaded server before it becomes an outage and getting time back from firefighting. Start a free trial and set your first alerts today.
Download 30-Day Trial

Product Details

Platforms
Web-Based
Types of Training
Training Docs
Live Training (Online)
Training Videos
Customer Support
Online Support

Mercury 2.5 Features and Options

Mercury 2.5 User Reviews

Write a Review
  • Previous
  • Next