FFree
InceptionClosed Weights

Mercury 2.5 Preview

inception/mercury-2.5-preview

Best Input Price

$0.04 / 1M

Mercury 2.5 Preview is Inception's latest and most intelligent diffusion language model. Instead of generating tokens strictly one at a time, it produces and refines multiple tokens in parallel, reaching up to 1,107 tokens per second on standard GPUs. It delivers a 10+ point intelligence gain over Mercury 2, with tunable reasoning, parallel tool calls, schema-aligned JSON output, and a 260K context window. It is built for latency-sensitive production work such as search agents, voice pipelines, customer support, rapid coding iteration, and coding subagents.

Context Window

260,000 tokens

Reasoning

Supported

Tool Calling

Supported

Released

2026-09-01

Inference Providers (1)

ProviderModel IDContextInput / 1MOutput / 1MAction
NanoGPTinception/mercury-2.5-preview260,000$0.04$0.15Docs ↗