Skip to content
AI Atlas
ModelActive

Mercury 2.5

Inception

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

quality50

Updated 4 h ago · first seen 11 Sept 2026

model_01M294WVMPVMQCDAFA2RZ20WT1

Context
260K tokens
T2 · 4 h ago
Released
8 Sept 2026
T2 · 4 h ago

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history

12 claims · 12 properties

Release daterelease_date1

Claim history for Release date
ValueValid from → toStatusSourceConfidenceExtractor
8 Sept 2026currentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Context windowcontext_length1

Claim history for Context window
ValueValid from → toStatusSourceConfidenceExtractor
260K tokenscurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Max outputmax_output_tokens1

Claim history for Max output
ValueValid from → toStatusSourceConfidenceExtractor
65.5K tokenscurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Modalitiesmodalities1

Claim history for Modalities
ValueValid from → toStatusSourceConfidenceExtractor
textcurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Input modalitiesmodalities_input1

Claim history for Input modalities
ValueValid from → toStatusSourceConfidenceExtractor
textcurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Output modalitiesmodalities_output1

Claim history for Output modalities
ValueValid from → toStatusSourceConfidenceExtractor
textcurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Descriptiondescription1

Claim history for Description
ValueValid from → toStatusSourceConfidenceExtractor
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...currentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Openrouter idopenrouter_id1

Claim history for Openrouter id
ValueValid from → toStatusSourceConfidenceExtractor
inception/mercury-2.5currentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Reasoningreasoning1

Claim history for Reasoning
ValueValid from → toStatusSourceConfidenceExtractor
YescurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Structured outputstructured_output1

Claim history for Structured output
ValueValid from → toStatusSourceConfidenceExtractor
YescurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Supported parameterssupported_parameters1

Claim history for Supported parameters
ValueValid from → toStatusSourceConfidenceExtractor
include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, stop, structured_outputs, temperature, tool_choice, toolscurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Tool callingtool_calling1

Claim history for Tool calling
ValueValid from → toStatusSourceConfidenceExtractor
YescurrentcurrentOpenRouter public model & pricing listingT2mediumdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →