Claude Opus 4.7
- Arena
- 1503
- Output (from)
- $25.00 / 1M
Last refreshed 2026-07-16. Next refresh: weekly.
The best LLMs for writing in 2026, ranked by human preference. Covers long-form essays, creative prose, and marketing copy — with pricing.
Looking for ad copy, email, social posts, localization, or brand voice? Use the marketing use-case matrix instead of treating general prose quality as the whole decision.
Verdict
Claude Opus 4.6 is the runner-up: 1503 vs 1501 on Arena.
Writing picks are for essays, drafts, and long-form prose. The target rubric is EQ-Bench Creative Writing v3 when coverage is available; the live fallback is Chatbot Arena, then MMLU.
| # | Model | Input $/1M | Output $/1M | |
|---|---|---|---|---|
| 1 | Claude Opus 4.7 ReasoningVisionTools Arena: 1503 | $5.00 | $25.00 | |
| 2 | Claude Opus 4.6 ReasoningVisionTools Arena: 1501 | $5.00 | $25.00 | |
| 3 | Gemini 3.1 Pro Preview PreviewVisionTools Arena: 1493 | $2.00 | $12.00 | |
| 4 | Muse Spark ReasoningVisionTools Arena: 1491 | — | — | |
| 5 | GPT-5.5 ReasoningVisionTools Arena: 1488 | $5.00 | $30.00 | |
| 6 | Gemini 3 Pro VisionTools Arena: 1486 | $1.25 | $5.00 | |
| 7 | GPT-5.4 ReasoningVisionTools Arena: 1479 | $2.50 | $15.00 | |
| 8 | ERNIE 5.1 Tools Arena: 1476 | $0.59 | $2.65 | |
| 9 | Qwen3.7-Max ReasoningTools Arena: 1475 | $1.25 | $3.75 | |
| 10 | GLM-5.1 ReasoningTools Arena: 1475 | $1.05 | $3.50 | |
| 11 | Gemini 3 Flash PreviewVisionTools Arena: 1467 | $0.50 | $3.00 | |
| 12 | Claude Opus 4.5 ReasoningVisionTools Arena: 1466 | $5.00 | $25.00 | |
| 13 | Grok 4.1 ReasoningVisionTools Arena: 1464 | — | — | |
| 14 | Claude Sonnet 4.6 ReasoningVisionTools Arena: 1459 | $3.00 | $15.00 | |
| 15 | DeepSeek V4 Pro ReasoningTools Arena: 1456 | $0.43 | $0.87 | |
| 16 | DeepSeek V4 Flash ReasoningTools Arena: 1437 | $0.09 | $0.18 | |
| 17 | Gemini 3.1 Flash-Lite VisionTools Arena: 1432 | $0.25 | $1.50 | |
| 18 | o3 ReasoningVisionTools Arena: 1412 | $2.00 | $8.00 | |
| 19 | Gemini 2.5 Pro ReasoningVisionTools Arena: 1398 | $1.25 | $10.00 | |
| 20 | DeepSeek R1 Reasoning Arena: 1372 | $0.10 | $0.30 |
Next seats in this ranking. Lines below are from each model's stored description in LLMReference seed data—spot-check the model page before relying on a capability claim.
Google DeepMind's most advanced reasoning Gemini model. Part of the Gemini 3 series with frontier-class intelligence, multimodal understanding, and 1M token context window.
1486
Arena
GPT-5.4 is OpenAI's flagship frontier reasoning model, released March 5, 2026. It incorporates advances from GPT-5.3-Codex for coding and agentic workflows, and adds 'Thinking' mode with editable reasoning plans. Key capabilities include computer use (navigating interfaces via Playwright), image understanding and generation integration, full-stack web app generation, tool calling, and deep research. Knowledge cutoff is August 31, 2025. Model ID: gpt-5.4.
1479
Arena
ERNIE 5.1 is Baidu's fifth-generation flagship language model, officially released May 9, 2026. Achieved via disaggregated fully-asynchronous reinforcement learning and scaled agentic post-training, it delivers leading performance at approximately 6% of the pre-training compute cost of comparable models — with roughly one-third the total parameters and half the active parameters of ERNIE 5.0. ERNIE 5.1 ranks #4 globally and #1 among Chinese models on the LMArena Search leaderboard (score: 1,223), with standout performance in legal reasoning, mathematics (AIME26: 99.6), and business domains. API model ID: ernie-5.1. Context: 128K tokens; max output: 65,536 tokens.
1476
Arena
Side-by-side comparison of the top picks by price, benchmark, and API access.
Claude Opus 4.7 is the current LLMReference top pick for writing. The verdict uses the stored category signal Arena: 1503. Output pricing starts at $25.00 per 1M tokens. Review the linked model and provider pages before production use because availability and pricing can change.
Claude Opus 4.7 leads Claude Opus 4.6 in the visible shortlist on Arena: 1503 versus 1501. The pricing cards show Claude Opus 4.7: output pricing starts at $25.00 per 1m tokens and Claude Opus 4.6: output pricing starts at $25.00 per 1m tokens.
LLMReference ranks LLMs for writing from stored model, benchmark, freshness, and pricing data. The current methodology summary is: Writing picks are for essays, drafts, and long-form prose. The target rubric is EQ-Bench Creative Writing v3 when coverage is available; the live fallback is Chatbot Arena, then MMLU.
The LLM rankings on this page are updated daily as new benchmark scores, provider availability, and pricing data are tracked. The "as of" date at the top of the page shows the most recent refresh.
The podium picks are driven by the primary benchmark signal for this category (shown in the Methodology section), filtered to non-deprecated models with confirmed API availability. In ties, we prefer the more recently released model.
Preview models appear in the "Watch list" section but are not in the main ranked podium unless the category explicitly allows it (e.g., /best/coding and /best/agents, where preview models often lead benchmarks).
Yes — use the Compare tool at llmreference.com/compare for a side-by-side breakdown of context window, pricing, benchmarks, and provider availability.
Pricing is tracked from provider documentation and updated regularly. It reflects the best available public data, not live API quotes — always verify before billing.