GPT-3.5 Turbo 16k on OpenRouter

GPT-3.5 · OpenAI

Serverless

Last refreshed 2026-06-15. Next refresh: weekly.

Why use GPT-3.5 Turbo 16k on OpenRouter?

OpenRouter offers GPT-3.5 Turbo 16k with pay-as-you-go pricing at $3.00/1M input tokens. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.

Compare GPT-3.5 Turbo 16k across 3 providers to find the best fit for your use case
Input / 1M
$3.00
Output / 1M
$4.00
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: openai/gpt-3.5-turbo-16k
Model ID
openai/gpt-3.5-turbo-16k

Request example

Curated snippets for this provider are not sourced yet. Use OpenRouter documentation with model ID openai/gpt-3.5-turbo-16k.

Gotchas

  • Use provider model ID "openai/gpt-3.5-turbo-16k", not the LLMReference slug "gpt-3.5-turbo-16k".

Compare GPT-3.5 Turbo 16k Across Providers

ProviderInput (per 1M)Output (per 1M)
Azure OpenAI$0.50$2.00
Salesforce Einstein Generative AI——
OpenRouter$3.00$4.00

Pricing

TypePrice (per 1M)
Input tokens$3.00
Output tokens$4.00

Capabilities

Structured Outputs

About GPT-3.5 Turbo 16k

GPT-3.5 Turbo 16k, developed by OpenAI, is an advanced language model featuring a significantly enhanced context window of 16,384 tokens—four times larger than its predecessor's 4,096 tokens 245. This extension allows it to process and comprehend extended texts, up to approximately 20 pages, in a single interaction, while maintaining the speed and efficiency of earlier versions 10. Although it is a chat-centric model not compatible with the completions endpoint 8, it remains highly effective for tasks requiring prolonged relevance and coherence through OpenAI's API 46.

Get Started

Model Specs

Released2023-06-13
Parameters20B
Context16k
ArchitectureDecoder Only
Knowledge cutoff2021-09

Related Models on OpenRouter