Mercury 2.5

mercury-2.5 · Inception

Mercury 2.5 is Inception's diffusion language model (dLLM), designed for text generation and reasoning tasks. Instead of sequential generation, it produces and refines multiple tokens in parallel to achieve rapid reasoning, supported by an expansive 260,000-token context window.

API Pricing

Input$0.2 / 1M tokens
Output$0.75 / 1M tokens
Cache read$0.02 / 1M tokens

Specifications

Context260K tokens
Modalitiestext

Frequently asked questions

What is Mercury 2.5?

Mercury 2.5 is Inception's diffusion language model (dLLM), designed for text generation and reasoning tasks. Instead of sequential generation, it produces and refines multiple tokens in parallel to achieve rapid reasoning, supported by an expansive 260,000-token context window.

What is the context length of Mercury 2.5?

Mercury 2.5 has a 260,000 token context window.

How much does Mercury 2.5 cost?

On AIHubMix, Mercury 2.5 costs $0.2 per million input tokens and $0.75 per million output tokens. Cached input reads are billed at $0.02 per million tokens.

What modalities does Mercury 2.5 support?

Mercury 2.5 accepts text input.

How do I call Mercury 2.5 via API?

Mercury 2.5 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to mercury-2.5 — no other code changes needed.

Who created Mercury 2.5?

Mercury 2.5 is developed by Inception. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Inception

See all Inception models →

Mercury 2.5 Preview

by Inception

Mercury 2.5 is the latest diffusion-based large language model (dLLM) released by…

$0.2 $0.04/1M in · $0.75 $0.15/1M out
80% off
260,000 tokens context

Use Mercury 2.5 via the AIHubMix unified API — one interface for every major LLM.