Inception: Mercury 2.5 Preview

Inception mercury-2-5-preview
Model Information
Slug mercury-2-5-preview
LLMs.txt View
Release Date August 31, 2026 New
Organization
Model Description
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.
Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving 1,107 tokens/sec on standard GPUs. It delivers a 10+ point jump in intelligence over Mercury 2, comparable quality to cost-optimized frontier models like GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5.
Mercury 2.5 supports tunable reasoning levels, parallel tool calls, and schema-aligned JSON output. It's built for production workloads where latency compounds: search agents, voice pipelines, and coding subagents.
Available at 9 Providers
Provider Type Model Name Original Model Input ($/1M) Output ($/1M) Free Actions
Nous Research
Nous Research
Inception: Mercury 2.5 Preview
inception/mercury-2.5-preview $0.032 $0.12
Cline
Cline
Code
Inception: Mercury 2.5 Preview
inception/mercury-2.5-preview $0.04 $0.15
OpenRouter
OpenRouter
Chat Code
Mercury 2.5 Preview
inception/mercury-2.5-preview $0.04 $0.15
Nano-GPT
Nano-GPT
Chat Code
Mercury 2.5 Preview
inception/mercury-2.5-preview $0.04 $0.15
Routeway
Routeway
Inception Labs: Mercury 2.5 Preview
mercury-2.5-preview $0.04 $0.15
Kilo Code
Kilo Code
Code
Inception: Mercury 2.5 Preview
inception/mercury-2.5-preview $0.20 $0.75
AIHubMix
AIHubMix
Chat Code
mercury-2.5-preview
mercury-2.5-preview $0.20 $0.75
Krater
Krater
Inception: Mercury 2.5 Preview
mercury-2-5-preview - -
Writingmate
Writingmate
Chat Code
Inception: Mercury 2.5 Preview
inception/mercury-2.5-preview - -