LLM Reference

Best Long Context LLMs (2026)

Last refreshed 2026-09-03. Next refresh: weekly.

LLMs with the largest context windows in 2026 — from 128K to 2M tokens. Ranked by window size with pricing and retrieval accuracy.

Release watch. Gemini 3.5 Pro is tracked with a 2M-token context window while it remains in preview.

Verdict

Use GPT-5.6 Sol for long-context work today.

GPT-5.6 Terra is the runner-up: 1.05m vs 1.05m on Context.

Researched 57d agoWhy this pickMethodology

How we rank

Long-context leaders sort by the widest shipped context window first, then recency — benchmarks matter, but literal window size is the gate for this category.

  1. EligibilityModels with ≥128k tokens of declared context and not deprecated.
  2. Primary rankingLarger parsed context wins; newer release on ties.
  3. Variant collapseWe keep one row per model family (`familySlug` + parameter tier). When headline scores tie within ±0.5 pt (±10 Elo on Chatbot Arena), we pick the canonical SKU by lowest tracked input price, then GA over preview or limited access, then newest `release`. A folded sibling within the benchmark noise band can show a "Tied within margin" chip on that score cell.
  4. PricingLowest tracked provider rates where available.
#ModelInput $/1MOutput $/1M
1RWKV-7 Goose 2.9B
2RWKV-7 Goose 1.5B
3RWKV-7 Goose 0.4B
4RWKV-7 Goose 0.1B
5RWKV-6 Finch 14B
6RWKV-6 Finch 7B
7RWKV-6 Finch 3B
8RWKV-6 Finch 1.6B
9LTM-2-mini
10Llama 4 Scout 17B-16E Instruct
Vision
$0.08$0.22
11LTM-1
12MiniMax-01
Vision
$0.20$1.10
13Gemini 3.5 Pro
PreviewReasoningVision
14Gemini 1.5 Pro 002
15Gemini 1.5 Pro Experimental 0827
16Gemini 1.5 Pro Experimental 0801
17Gemini 1.5 Pro
$1.25$5.00
18GPT-6 Astra
RestrictedReasoningVisionTools
19GPT-5.6 Sol
ReasoningVisionTools
$5.00$30.00
20GPT-5.6 Terra
ReasoningVisionTools
$2.50$15.00
21GPT-5.6 Luna
ReasoningVisionTools
$1.00$6.00
22GPT-5.5
ReasoningVisionTools
$5.00$30.00
23GPT-5.5 Pro
ReasoningVisionTools
$30.00$180.00
24GPT-5.4
ReasoningVisionTools
$2.50$15.00
25GPT-5.4 Pro
ReasoningVisionTools
$30.00$180.00
26Muse Spark 1.3
ReasoningVisionTools
$1.25$4.25
27Muse Spark 1.3 Contributor
ReasoningVisionTools
$0.10$0.20
28Gemini 3.8 Flash
ReasoningVisionTools
$0.75$3.75
29Spark X2.5 4B
ReasoningTools
30Spark X2.5 1.7B
ReasoningTools
31Hunyuan Hy4 Preview
PreviewReasoningTools
$0.83$2.50
32Gemini Omni 1.1 Flash
Vision
$1.50$17.50
33Gemini 3.7 Flash
ReasoningVisionTools
$0.75$3.75
34Muse Spark 1.2
PreviewReasoningVisionTools
$1.25$4.25
35Gemini 3.6 Flash
ReasoningVisionTools
$1.50$7.50
36Gemini 3.5 Flash-Lite
ReasoningVisionTools
$0.30$2.50
37Kimi K3
ReasoningVisionTools
$3.00$15.00
38Gemini 3.5 Flash
ReasoningVisionTools
$1.50$9.00
39Antigravity Agent
PreviewReasoningVision
40Gemini 3.1 Flash-Lite
VisionTools
$0.25$1.50

Honorable mentions

Next seats in this ranking. Lines below are from each model's stored description in LLMReference seed data—spot-check the model page before relying on a capability claim.
  • #4GPT-5.5

    GPT-5.5 is OpenAI's fully retrained agentic model, released April 23, 2026. Optimised for agentic coding, computer use, knowledge work, and early scientific research. Achieves 82.7% on Terminal-Bench 2.0 (Codex CLI scaffold), 84.9% on GDPval, 58.6% on SWE-Bench Pro, 93.6% on GPQA Diamond, and 82.6% on SWE-Bench Verified (Vals.ai independent harness). Knowledge cutoff December 2025. Supports reasoning effort levels (none/low/medium/high/xhigh). Context window 1,050,000 tokens with a long-context surcharge above 272K tokens. Model ID: gpt-5.5.

    1.05m

    Context

  • #5GPT-5.5 Pro

    GPT-5.5 Pro is OpenAI's premium extra-compute deployment of GPT-5.5, released April 23, 2026. It uses the same underlying weights as GPT-5.5 standard with additional parallel test-time compute for harder tasks. Supports text and image inputs, reasoning effort control, tool use, structured outputs, code execution, a 1,050,000-token context window, and 128K max output. Key datapack rows: Terminal-Bench 2.1 78.2%, SWE-bench Pro 58.6%, GPQA Diamond 93.6%, ARC-AGI-2 high effort 83.3%, BrowseComp Pro compute 90.1%, and FrontierMath Tier 4 39.6%. Official pricing is $30/M input, $180/M output, $10/M batch input, and $45/M batch output; native cached input discount is not listed.

    1.05m

    Context

  • #6GPT-5.4

    GPT-5.4 is OpenAI's flagship frontier reasoning model, released March 5, 2026. It incorporates advances from GPT-5.3-Codex for coding and agentic workflows, and adds 'Thinking' mode with editable reasoning plans. Key capabilities include computer use (navigating interfaces via Playwright), image understanding and generation integration, full-stack web app generation, tool calling, and deep research. Knowledge cutoff is August 31, 2025. Model ID: gpt-5.4.

    1.05m

    Context