Back to all models
inception

Inception: Mercury 2.5 Preview

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5-preview

Context Size

260K

Input Price

6,000 pts/M

Output Price

22,500 pts/M


Architecture

Text

Supported Parameters

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetools

Details

TokenizerOther
Max Completion65,536 tokens
Provider Context260K tokens
ModeratedNo
CreatedAug 31, 2026