Inception: Mercury 2.5 Preview

Text input Text output
Author's Description

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Key Specifications
Cost
$$$
Context
260K
Released
Aug 31, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Stop Temperature Tool Choice Response Format Include Reasoning Reasoning Tools Max Tokens Structured Outputs
Features

This model supports the following features:

Reasoning Tools Response Format Structured Outputs
Performance Summary

Inception: Mercury 2.5 Preview, created on August 31, 2026, is positioned as a groundbreaking diffusion LLM (dLLM) that generates and refines tokens in parallel, aiming for unparalleled speed. This model consistently ranks among the fastest and most competitively priced models across eight benchmarks, achieving an "Infinityth percentile" in both speed and price. Its reliability is exceptional, boasting a 100% success rate across all benchmarks, indicating minimal technical failures. Performance across benchmarks reveals a model with significant strengths. Mercury 2.5 achieved perfect 100.0% accuracy in General Knowledge, notably being the most accurate model at its price point and among models of comparable speed. It also demonstrated strong performance in Coding (95.7% accuracy, 91st percentile), Reasoning (98.0% accuracy, 91st percentile), and Mathematics (97.0% accuracy, 97th percentile), showcasing robust analytical and problem-solving capabilities. A notable weakness is its 0.0% accuracy in Instruction Following, suggesting a critical area for improvement. Hallucinations were relatively low at 98.0% accuracy, indicating a good ability to acknowledge uncertainty.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.04
Completion $0.15
Input Cache Read $0.004

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Inception
Inception | inception/mercury-2.5-preview-20260831 260K $0.04 / 1M tokens $0.15 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by inception