GPT-4 Turbo
GPT-4 Turbo is a legacy integration reference; evaluate GPT-4.1 before starting new work.
- Family
- GPT-4
- Released
- 2024-04-09
- Context
- 128k
- Parameters
- 1.76T (8x222B MoE)*
- Architecture
- Mixture of Experts
- Knowledge cutoff
- 2023-12
- Specialization
- general
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
- Training
- Fine-tuned
Cheapest of 6 routes · Replicate API
This model is deprecated. OpenAI recommends switching to GPT-4.1.
About
OpenAI's high-performance variant of GPT-4 with 128K context window and improved reasoning. Widely used for production applications.
Provider price ladder
Compare all 6Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Batch in / out | Route |
|---|---|---|---|---|
| Replicate API | $5.00 | $15.00 | - | Serverless |
| Azure OpenAI | $10.00 | $30.00 | - | Serverless |
| OpenAI API | $10.00 | $30.00 | $5.00 / $15.00 | Serverless |
| OpenRouter | $10.00 | $30.00 | - | Serverless |
Available via routers & gateways(16)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Portkey
GatewayProduction AI gateway routing to 1,600+ LLMs with failover, load balancing, semantic caching, and guardrails; Apache 2.0 core is fully self-hostable with the complete feature set.
AIRouter
RouterCommercial LLM router that analyzes incoming requests and routes to the optimal model for cost/quality/latency via a drop-in OpenAI-compatible API, with a privacy-preserving embedding mode that avoids sending prompt content.
Azure AI Foundry Model Router
RouterMicrosoft Azure AI Foundry's native model router that uses a trained ML model to route each prompt in real time to the optimal Azure-hosted model, with Balanced/Cost/Quality mode selection and automatic failover.
Helicone
GatewayObservability-first AI gateway with routing, caching, rate limiting, and request tracing; Apache 2.0 open-source core with a managed hosted tier for logging and analytics.
Capabilities
Benchmark peer barsfor Classification
Benchmark scores(3)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Massive Multitask Language Understanding | 86.5 | 5-shotObserved 2026-03-07 | — | Source |
| Chatbot Arena | 1242.0 | —Observed 2026-04-15 | — | Source |
| MMLU PRO | 69.4 | —Observed 2026-04-14 | — | Source |
Migration checks
No linked migration route is available for this model yet.
API versions
gpt-4-turbo-2024-04-09gpt-4-turboCompare GPT-4 Turbo with other models
Comparison and alternatives
Browse all comparisons →Show all 2 popular comparisonssorted by 7-day search impressions
Cheapest of 6 routes · Replicate API