LLM Reference

Best Long Context LLMs (2026)

Last refreshed 2026-08-28. Next refresh: weekly.

LLMs with the largest context windows in 2026 — from 128K to 2M tokens. Ranked by window size with pricing and retrieval accuracy.

Release watch. Gemini 3.5 Pro is tracked with a 2M-token context window while it remains in preview.

Verdict

Use GPT-5.6 Sol for long-context work today.

GPT-5.6 Terra is the runner-up: 1.05m vs 1.05m on Context.

Researched 50d agoWhy this pickMethodology

How we rank

Long-context leaders sort by the widest shipped context window first, then recency — benchmarks matter, but literal window size is the gate for this category.

  1. EligibilityModels with ≥128k tokens of declared context and not deprecated.
  2. Primary rankingLarger parsed context wins; newer release on ties.
  3. Variant collapseWe keep one row per model family (`familySlug` + parameter tier). When headline scores tie within ±0.5 pt (±10 Elo on Chatbot Arena), we pick the canonical SKU by lowest tracked input price, then GA over preview or limited access, then newest `release`. A folded sibling within the benchmark noise band can show a "Tied within margin" chip on that score cell.
  4. PricingLowest tracked provider rates where available.
#ModelInput $/1MOutput $/1M
1RWKV-7 Goose 2.9B
2RWKV-7 Goose 1.5B
3RWKV-7 Goose 0.4B
4RWKV-7 Goose 0.1B
5RWKV-6 Finch 14B
6RWKV-6 Finch 7B
7RWKV-6 Finch 3B
8RWKV-6 Finch 1.6B
9LTM-2-mini
10Llama 4 Scout 17B-16E Instruct
Vision
$0.08$0.22
11LTM-1
12MiniMax-01
Vision
$0.20$1.10
13Gemini 3.5 Pro
PreviewReasoningVision
14Gemini 1.5 Pro 002
15Gemini 1.5 Pro Experimental 0827
16Gemini 1.5 Pro Experimental 0801
17Gemini 1.5 Pro
$1.25$5.00
18GPT-5.6 Sol
ReasoningVisionTools
$5.00$30.00
19GPT-5.6 Terra
ReasoningVisionTools
$2.50$15.00
20GPT-5.6 Luna
ReasoningVisionTools
$1.00$6.00
21GPT-5.5
ReasoningVisionTools
$5.00$30.00
22GPT-5.5 Pro
ReasoningVisionTools
$30.00$180.00
23GPT-5.4
ReasoningVisionTools
$2.50$15.00
24GPT-5.4 Pro
ReasoningVisionTools
$30.00$180.00
25Gemini Omni 1.1 Flash
Vision
$1.50$17.50
26Gemini 3.7 Flash
ReasoningVisionTools
$0.75$3.75
27Muse Spark 1.2
PreviewReasoningVisionTools
$1.25$4.25
28Gemini 3.6 Flash
ReasoningVisionTools
$1.50$7.50
29Gemini 3.5 Flash-Lite
ReasoningVisionTools
$0.30$2.50
30Kimi K3
ReasoningVisionTools
$3.00$15.00
31Gemini 3.5 Flash
ReasoningVisionTools
$1.50$9.00
32Antigravity Agent
PreviewReasoningVision
33Gemini 3.1 Flash-Lite
VisionTools
$0.25$1.50
34Xiaomi MiMo-V2.5-Pro
Tools
$0.43$0.87
35Xiaomi MiMo-V2.5
ReasoningVisionTools
$0.14$0.28
36Nemotron-Cascade-2-30B-A3B
37MiMo-V2-Pro
$1.00$3.00
38Nemotron 3 Super-120B-A12B
$0.09$0.45
39Gemini 2.5 Pro Computer Use Preview
PreviewVisionTools
$1.25$10.00
40Llama 3 70B Gradient 1048K

Honorable mentions

Next seats in this ranking. Lines below are from each model's stored description in LLMReference seed data—spot-check the model page before relying on a capability claim.
  • #4GPT-5.5

    GPT-5.5 is OpenAI's fully retrained agentic model, released April 23, 2026. Optimised for agentic coding, computer use, knowledge work, and early scientific research. Achieves 82.7% on Terminal-Bench 2.0 (Codex CLI scaffold), 84.9% on GDPval, 58.6% on SWE-Bench Pro, 93.6% on GPQA Diamond, and 82.6% on SWE-Bench Verified (Vals.ai independent harness). Knowledge cutoff December 2025. Supports reasoning effort levels (none/low/medium/high/xhigh). Context window 1,050,000 tokens with a long-context surcharge above 272K tokens. Model ID: gpt-5.5.

    1.05m

    Context

  • #5GPT-5.5 Pro

    GPT-5.5 Pro is OpenAI's premium extra-compute deployment of GPT-5.5, released April 23, 2026. It uses the same underlying weights as GPT-5.5 standard with additional parallel test-time compute for harder tasks. Supports text and image inputs, reasoning effort control, tool use, structured outputs, code execution, a 1,050,000-token context window, and 128K max output. Key datapack rows: Terminal-Bench 2.1 78.2%, SWE-bench Pro 58.6%, GPQA Diamond 93.6%, ARC-AGI-2 high effort 83.3%, BrowseComp Pro compute 90.1%, and FrontierMath Tier 4 39.6%. Official pricing is $30/M input, $180/M output, $10/M batch input, and $45/M batch output; native cached input discount is not listed.

    1.05m

    Context

  • #6GPT-5.4

    GPT-5.4 is OpenAI's flagship frontier reasoning model, released March 5, 2026. It incorporates advances from GPT-5.3-Codex for coding and agentic workflows, and adds 'Thinking' mode with editable reasoning plans. Key capabilities include computer use (navigating interfaces via Playwright), image understanding and generation integration, full-stack web app generation, tool calling, and deep research. Knowledge cutoff is August 31, 2025. Model ID: gpt-5.4.

    1.05m

    Context