Mercury 2
Mercury 2 is worth evaluating for rag, long context, and classification when its provider route and context window match the workload.
Use it for
- Teams evaluating rag, long context, and classification
- Workloads that can use a 131k context window
- Buyers comparing 2 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
- Family
- Mercury
- Released
- 2026-02-20
- Context
- 131k
- Openness
- Proprietary
- Weights
- Not released
- Code
- Unknown
Inception Labs is an AI company building diffusion-based large language models (dLLMs), a fundamentally different archit
Cheapest of 2 routes · OpenRouter
About
Inception Labs' Mercury 2 is a commercial-scale diffusion-based language model (dLLM) released February 2026. Unlike autoregressive transformers, Mercury uses a diffusion architecture for faster inference. Designed for code generation, reasoning, and analysis tasks. Proprietary, available via the Inception API.
Top use-case fit
RAG
Included by capability and metadata signals in the decision map.
Long context
Included by capability and metadata signals in the decision map.
Classification
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare all 2Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| OpenRouter | $0.250 | $0.750 | - | Serverless |
| Vercel AI Gateway | $0.250 | $0.750 | read $0.025 | Serverless |
Capabilities
Benchmark peer barsfor RAG
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Inception Labs is an AI company building diffusion-based large language models (dLLMs), a fundamentally different archit
Cheapest of 2 routes · OpenRouter