LLM Reference

Mistral Large 2 vs Qwen3.5-122B-A10B

Mistral Large 2 (2025) and Qwen3.5-122B-A10B (2026) are frontier reasoning models from MistralAI and Alibaba. Mistral Large 2 ships a 128k-token context window, while Qwen3.5-122B-A10B ships a 262k-token context window. On MMLU PRO, Qwen3.5-122B-A10B leads by 17 pts. On pricing, Qwen3.5-122B-A10B costs $0.26/1M input tokens versus $0.48/1M for the alternative. This comparison covers specs, pricing, API access, capabilities, benchmarks, input and output token costs, and production fit for coding and agent workloads.

Qwen3.5-122B-A10B is ~85% cheaper at $0.26/1M; pay for Mistral Large 2 only for vision-heavy evaluation.

Decision scorecard

Local evidence first
SignalMistral Large 2Qwen3.5-122B-A10B
Best formultimodal apps, tool-calling agents, and provider-routed productionreasoning-heavy apps, multimodal apps, and tool-calling agents
Decision fitCoding, RAG, and AgentsCoding, RAG, and Agents
Context window128k262k
Cheapest output$2.40/1M tokens$2.08/1M tokens
Provider routes3 tracked3 tracked
Shared benchmarks1 sharedMMLU PRO leader

Decision tradeoffs

Choose Mistral Large 2 when...
  • Local decision data tags Mistral Large 2 for Coding, RAG, and Agents.
Choose Qwen3.5-122B-A10B when...
  • Qwen3.5-122B-A10B holds a shared-benchmark lead on MMLU PRO, ahead by 17 points.
  • Qwen3.5-122B-A10B has the larger context window for long prompts, retrieval packs, or transcript analysis.
  • Qwen3.5-122B-A10B has the lower cheapest tracked output price at $2.08/1M tokens.
  • Qwen3.5-122B-A10B uniquely exposes Reasoning in local model data.
  • Local decision data tags Qwen3.5-122B-A10B for Coding, RAG, and Agents.

Monthly cost at traffic

Estimate token spend from the cheapest tracked input and output route or tier on this page.

Lower estimate Qwen3.5-122B-A10B

Mistral Large 2

$984

Cheapest tracked route/tier: AWS Bedrock

Qwen3.5-122B-A10B

$728

Cheapest tracked route/tier: OpenRouter

Estimated monthly gap: $256. Batch, cache, alternate speed tiers, and negotiated pricing are excluded from this local estimate.

Switch friction

Mistral Large 2 -> Qwen3.5-122B-A10B
  • No overlapping tracked provider route is sourced for Mistral Large 2 and Qwen3.5-122B-A10B; plan for SDK, billing, or endpoint changes.
  • Qwen3.5-122B-A10B is $0.32/1M tokens lower on cheapest tracked output pricing before cache, batch, or negotiated discounts.
  • Qwen3.5-122B-A10B adds Reasoning in local capability data.
Qwen3.5-122B-A10B -> Mistral Large 2
  • No overlapping tracked provider route is sourced for Qwen3.5-122B-A10B and Mistral Large 2; plan for SDK, billing, or endpoint changes.
  • Mistral Large 2 is $0.32/1M tokens higher on cheapest tracked output pricing, so quality gains need to justify the spend.
  • Check replacement coverage for Reasoning before moving production traffic.

Specs

Specification
Released2025-11-252026-02-24
Context window128k262k
Parameters123B122B
ArchitectureDecoder OnlyMixture of Experts
LicenseMistral LicenseApache 2.0OSI-approved
OpennessOpen weightsOpen source
WeightsUnknownAvailable
CodeUnknownUnknown
Commercial useCommercial use: non-commercialCommercial use: permitted
Knowledge cutoff2025-07-

Pricing and availability

Pricing attributeMistral Large 2Qwen3.5-122B-A10B
Input price$0.48/1M tokens$0.26/1M tokens
Output price$2.40/1M tokens$2.08/1M tokens
Providers

Capabilities

CapabilityMistral Large 2Qwen3.5-122B-A10B
VisionYesYes
MultimodalYesYes
ReasoningNoYes
JSON / Tool useYesYes
Structured outputsYesYes
Code executionNoNo
IDE integrationNoNo
Computer useNoNo
Parallel agentsNoNo

Benchmarks

BenchmarkMistral Large 2Qwen3.5-122B-A10B
MMLU PRO69.786.7

Continue comparing

Last reviewed: 2026-06-29. Data sourced from public model cards and provider documentation.