Back to all models
inception
Inception: Mercury 2.5 Preview
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
inception/mercury-2.5-preview
Context Size
260K
Input Price
6,000 pts/M
Output Price
22,500 pts/M
Architecture
Text
Supported Parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetools
Details
TokenizerOther
Max Completion65,536 tokens
Provider Context260K tokens
ModeratedNo
CreatedAug 31, 2026