inception

Inception: Mercury 2.5

inception/mercury-2.5

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Prices updated: 10/7/2026

Input
4.4
credits / 1M tokens
Output
17
credits / 1M tokens
Context
260k
tokens
Messages per $1 USD
≈ 7,900
typical message: 1k tokens in + 500 out

Prices in credits per million tokens. 1 credit = $0.01 USD. Final price, billed in MXN at the exchange rate on payment day + VAT (CFDI). Subject to change without notice.

Additional prices

Cache (read): 0.4 /1M

Spec sheet

Inputs
text
Outputs
text
Max output tokens
65,536
Knowledge cutoff
—
Tokenizer
Other
Capabilities
Text / LLM
Tools
JSON
Reasoning

Use it in your code

One OpenAI-compatible endpoint. Switch models by changing one line.

curl https://api.deployalo.com/ai/gateway/v1/chat/completions \
  -H "Authorization: Bearer $DEPLOYALO_AI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"inception/mercury-2.5","messages":[{"role":"user","content":"Hola"}]}'

Ready to use it?

Top up a credit pack, create your API key with a spend cap and start in minutes.