GPT-4o
- Family
- GPT-4o
- Released
- 2024-05-13
- Context
- 128k
- Max output
- 16,384
- Architecture
- Decoder Only
- Knowledge cutoff
- 2023-10
- Specialization
- general
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
- Training
- Pretrained
Cheapest of 5 routes · OpenAI API · cache read $1.25
About
OpenAI GPT-4o: Flagship multimodal model with vision, function calling, and broad capability. $2.50/M input, $10/M output.
Provider price ladder
Compare all 5Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| OpenAI API | $2.50 | $10.00 | read $1.25 | Serverless |
| OpenRouter | $2.50 | $10.00 | - | Serverless |
| Replicate API | $2.50 | $10.00 | - | Serverless |
| Vercel AI Gateway | $2.50 | $10.00 | read $1.25 | Serverless |
Available via routers & gateways(16)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Portkey
GatewayProduction AI gateway routing to 1,600+ LLMs with failover, load balancing, semantic caching, and guardrails; Apache 2.0 core is fully self-hostable with the complete feature set.
AIRouter
RouterCommercial LLM router that analyzes incoming requests and routes to the optimal model for cost/quality/latency via a drop-in OpenAI-compatible API, with a privacy-preserving embedding mode that avoids sending prompt content.
Azure AI Foundry Model Router
RouterMicrosoft Azure AI Foundry's native model router that uses a trained ML model to route each prompt in real time to the optimal Azure-hosted model, with Balanced/Cost/Quality mode selection and automatic failover.
Helicone
GatewayObservability-first AI gateway with routing, caching, rate limiting, and request tracing; Apache 2.0 open-source core with a managed hosted tier for logging and analytics.
Capabilities
Benchmark peer barsfor Vision
Benchmark scores(4)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Chatbot Arena | 1315.0 | —Observed 2026-04-15 | — | Source |
| Google-Proof Q&A | 50.6 | diamondObserved 2026-04-15 | — | Source |
| Massive Multi-discipline Multimodal Understanding | 69.1 | —Observed 2026-04-15 | — | Source |
| MMMU Pro | 64.7 | standard 4-option (original paper harness), 0513 checkpointObserved 2024-09-04 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(1)
Compare GPT-4o with other models
Cheapest of 5 routes · OpenAI API · cache read $1.25