LLM Reference

Best LLM for translation

Last refreshed 2026-09-03. Next refresh: weekly.

Compare multilingual LLMs and dedicated translation models for text, document, and live speech translation. Ranked by context length and general-language benchmarks until translation-specific leaderboard rows land in seed data.

Verdict

Use GPT-5.6 Sol for this task today.

GPT-5.6 Terra is the runner-up: Current leader vs No. 2 on Pick.

Researched 57d agoWhy this pickMethodology

How we rank

Translation picks are methodology-forward until dedicated translation benchmark rows land in seed data. We surface tagged translation models, Qwen-MT specialists, and long-context chat models teams commonly route for multilingual work.

  1. EligibilityModels with translation use-case tags, translation specialization, Qwen-MT family membership, or ≥128K context on general chat-completion routes (excluding embeddings and media-only specialists).
  2. Primary rankingDeclared context window (wider first), then MMLU when present, then newer release. Realtime speech translation models appear in the table but are labeled separately from text-generation LLMs.
  3. Benchmark gapNo translation-quality leaderboard is standardized in seed yet — do not treat this page as a definitive translation-quality ranking until WMT or vendor-neutral multilingual rows are sourced.
  4. Pricing columnLowest tracked provider input/output where a public rate card exists.
#ModelInput $/1MOutput $/1M
1RWKV-7 Goose 2.9B
2RWKV-7 Goose 1.5B
3RWKV-7 Goose 0.4B
4RWKV-7 Goose 0.1B
5RWKV-6 Finch 14B
6RWKV-6 Finch 7B
7RWKV-6 Finch 3B
8RWKV-6 Finch 1.6B
9LTM-2-mini
10Llama 4 Scout 17B-16E Instruct
Vision
$0.08$0.22
11LTM-1
12MiniMax-01
Vision
$0.20$1.10
13Gemini 3.5 Pro
PreviewReasoningVision
14Gemini 1.5 Pro 002
15Gemini 1.5 Pro Experimental 0827
16Gemini 1.5 Pro Experimental 0801
17Gemini 1.5 Pro
$1.25$5.00
18GPT-5.5
ReasoningVisionTools
$5.00$30.00
19GPT-6 Astra
RestrictedReasoningVisionTools
20GPT-5.6 Sol
ReasoningVisionTools
$5.00$30.00

Honorable mentions

Next seats in this ranking. Lines below are from each model's stored description in LLMReference seed data—spot-check the model page before relying on a capability claim.
  • #4GPT-5.5 Pro

    GPT-5.5 Pro is OpenAI's premium extra-compute deployment of GPT-5.5, released April 23, 2026. It uses the same underlying weights as GPT-5.5 standard with additional parallel test-time compute for harder tasks. Supports text and image inputs, reasoning effort control, tool use, structured outputs, code execution, a 1,050,000-token context window, and 128K max output. Key datapack rows: Terminal-Bench 2.1 78.2%, SWE-bench Pro 58.6%, GPQA Diamond 93.6%, ARC-AGI-2 high effort 83.3%, BrowseComp Pro compute 90.1%, and FrontierMath Tier 4 39.6%. Official pricing is $30/M input, $180/M output, $10/M batch input, and $45/M batch output; native cached input discount is not listed.

    See leaderboard

    Rank

  • #5GPT-5.4

    GPT-5.4 is OpenAI's flagship frontier reasoning model, released March 5, 2026. It incorporates advances from GPT-5.3-Codex for coding and agentic workflows, and adds 'Thinking' mode with editable reasoning plans. Key capabilities include computer use (navigating interfaces via Playwright), image understanding and generation integration, full-stack web app generation, tool calling, and deep research. Knowledge cutoff is August 31, 2025. Model ID: gpt-5.4.

    See leaderboard

    Rank

  • #6GPT-5.4 Pro

    Premium extended-reasoning GPT-5.4 variant producing smarter and more precise responses. Replacement for o3-deep-research and o4-mini-deep-research. No prompt caching discount.

    See leaderboard

    Rank