LLM Reference

Mercury Models by Inception Labs

1 model2026Up to 131k ctxFrom $0.25/1M input

Last refreshed 2026-06-29. Next refresh: weekly.

Details

ResearcherInception Labs
Models1
Released2026
Max context131k

Capabilities

Structured OutputsAll models

Links

Website

About

Inception Labs' Mercury series of diffusion-based large language models (dLLMs). Mercury uses a fundamentally different architecture from autoregressive models, enabling faster generation speeds. Mercury 2 is a commercial-scale reasoning model optimized for code and analysis tasks.

Decision facts

Best fit
structured outputscoding
Capability starting point
Mercury 2 with 131k context and structured outputs
Lowest tracked input
Mercury 2 · $0.25/1M · OpenRouter
Closest related family
Claude 3

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view
Mercury 2Current

Use when the workload needs 131k context and structured outputs.

2026-02131k contextstructured outputs

Release Timeline

1 release group
2026-02
1 current
Mercury 2
131k contextstructured outputs
Current

Specifications(1 models)

Mercury model specifications comparison
ModelReleasedContextStructured Outputs
Mercury 22026-02131kYes

Available From(2 providers)

Pricing

Mercury model pricing by provider
ModelProviderInput / 1MOutput / 1MType
Mercury 2OpenRouter$0.25$0.75Serverless
Mercury 2Vercel AI Gateway$0.25$0.75Serverless