← model directory

inception

Inception: Mercury 2.5

inception/mercury-2.5

⚙ tool calling

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

specs & pricing
Type
text
Provider
inception
Model ID
inception/mercury-2.5
Capabilities
tools, reasoning
Context window
260K tokens
Input price
$0.044 / 1M tokens
Output price
$0.17 / 1M tokens

Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.