Inception model page

Inception: Mercury 2.5

inception/mercury-2.5

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

TextStructured outputTool callingReasoningLong context
Input / 1M
$0.05
Output / 1M
$0.20
Context
260K
Markup
30%

Use this model through UAI

A dedicated public landing page for this synced catalog row.

Provider
Inception

Synced from the active catalog

Context window
260K tokens

Validated for gateway requests

UAI input
$0.05

Charged from your UAI balance

UAI output
$0.20

Same API key and billing flow

This page exists for direct linking, onboarding, and SEO. The same model row powers /v1/models, the public catalog, the app pricing screen, and request validation for /v1/chat/completions.

Why use this model through UAI

The same model route, but with one account, one balance, and one API shape.

Reasoning workloads

Use Inception: Mercury 2.5 for multi-step prompts, planning, and deeper analysis while keeping the same OpenAI-compatible UAI endpoint.

Long-context inputs

Inception: Mercury 2.5 exposes 260K tokens, which makes it a good fit for large documents, transcripts, and retrieval-heavy chat flows.

Agent integrations

Tool and function-style parameters stay available through the same UAI request surface, so agent stacks do not need a separate provider integration.

Pricing and access

Current synced numbers for this exact model row.

UAI input / 1M tokens
$0.05
UAI output / 1M tokens
$0.20
Inception input / 1M
$0.04
Inception output / 1M
$0.15
Context window
260,000
Current markup
30%

Need another option? Browse the full models catalog or compare rows in the signed-in pricing table.

Quickstart

Use this model with your UAI key over the standard OpenAI-compatible route.

Create a key in API Keys, then send requests to https://uai.sh/v1/chat/completions.

curl
curl https://uai.sh/v1/chat/completions \
  -H "Authorization: Bearer uai-YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "inception/mercury-2.5",
    "messages": [{"role": "user", "content": "Say hi in one short sentence."}],
    "stream": false
  }'
Python (OpenAI SDK)
from openai import OpenAI

client = OpenAI(base_url="https://uai.sh/v1", api_key="uai-YOUR_API_KEY")
response = client.chat.completions.create(
    model="inception/mercury-2.5",
    messages=[{"role": "user", "content": "Say hi in one short sentence."}],
    stream=False
)
print(response.choices[0].message.content)
Node.js (OpenAI SDK)
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://uai.sh/v1",
  apiKey: "uai-YOUR_API_KEY"
});

const response = await client.chat.completions.create({
  model: "inception/mercury-2.5",
  messages: [{ role: "user", content: "Say hi in one short sentence." }],
  stream: false
});

console.log(response.choices[0].message.content);

Related models

Other active rows to compare without leaving the public catalog.

deepseek
DeepSeek: DeepSeek V4 Flash 0731
deepseek/deepseek-v4-flash-0731
Input / 1M$0.05
Context1.3M tokens
tencent
Tencent: Hy-MT2-1.8B
tencent/hy-mt2-1.8b
Input / 1M$0.06
Context8K tokens
inference-net
Inference.net: Schematron V2 Small
inference-net/schematron-v2-small
Input / 1M$0.07
Context128K tokens