Using GPT-3.5 Turbo 16k on Azure OpenAI

Implementation guide · GPT-3.5 · OpenAI

Serverless

Azure OpenAI exposes GPT-3.5 Turbo 16k through model ID gpt-3.5-turbo-16k. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-04-27. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Azure OpenAI and generate an API key.
  2. 2
    Use the Azure OpenAI SDK or REST API to call gpt-3.5-turbo-16k — see the documentation for request format.
  3. 3
    You'll be billed $0.50/1M input, $2.00/1M output tokens. See full pricing.

Code Examples

Install
pip install openai
API key
AZURE_OPENAI_API_KEY
Model ID
gpt-3.5-turbo-16k

gpt-3.5-turbo-16k is your Azure deployment name, not the underlying model name. Deployment names are set when you deploy a model in Azure AI Foundry / Azure OpenAI Studio.

import os
from openai import AzureOpenAI

client = AzureOpenAI(
    azure_endpoint=os.environ["AZURE_OPENAI_ENDPOINT"],  # e.g. https://{resource}.openai.azure.com
    api_key=os.environ["AZURE_OPENAI_API_KEY"],
    api_version="2024-02-01"
)
response = client.chat.completions.create(
    model="gpt-3.5-turbo-16k",  # your deployment name
    messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)

Pricing on Azure OpenAI

TypePrice (per 1M)
Input tokens$0.50
Output tokens$2.00

Capabilities

Structured Outputs

About GPT-3.5 Turbo 16k

GPT-3.5 Turbo 16k, developed by OpenAI, is an advanced language model featuring a significantly enhanced context window of 16,384 tokens—four times larger than its predecessor's 4,096 tokens 245. This extension allows it to process and comprehend extended texts, up to approximately 20 pages, in a single interaction, while maintaining the speed and efficiency of earlier versions 10. Although it is a chat-centric model not compatible with the completions endpoint 8, it remains highly effective for tasks requiring prolonged relevance and coherence through OpenAI's API 46.

Model Specs

Released2023-06-13
Parameters20B
Context16k
ArchitectureDecoder Only
Knowledge cutoff2021-09

Provider

Azure OpenAI

Microsoft

Redmond, Washington, United States