Mercury 2.5on Krater.ai

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving

Released September 8, 2026

Specifications

Mercury 2.5 at a glance

Context

260K tokens

Max output

65.5K tokens

Input

Text

Output

Text

Input price

$0.04 / 1M

Output price

$0.15 / 1M

Tool callingStructured outputsReasoning support

Capabilities

What it is good at

Catalog capabilities

  • Generates text responses
  • Supports tool calling
  • Supports structured outputs

Good for

Long documentsUse with toolsStructured responsesDeep analysis

FAQ

Questions about Mercury 2.5

Clear answers based on the model catalog and Krater access.

Mercury 2.5 supports a context window of 260,000 tokens.

Use Mercury 2.5 in your API

Connect with Krater's unified API.

API access

Ready to use Mercury 2.5?

Access leading models in one focused workspace.