AWS Bedrock exposes GPT-6 Luna through model ID openai.gpt-6-luna. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-09-23. Next refresh: weekly.
Quick Start
- 1
- 2Use the AWS Bedrock SDK or REST API to call
openai.gpt-6-luna— see the documentation for request format. - 3
Code Examples
pip install boto3AWS_ACCESS_KEY_IDopenai.gpt-6-lunaUse Amazon Bedrock model IDs, e.g. "anthropic.claude-3-opus-20240229-v1:0" for on-demand, or cross-region inference profile IDs like "us.anthropic.claude-opus-4-7-20251101-v1:0". These differ from the public model slug.
import boto3
# Reads AWS_ACCESS_KEY_ID, AWS_SECRET_ACCESS_KEY, AWS_DEFAULT_REGION from env
client = boto3.client("bedrock-runtime", region_name="us-east-1")
response = client.converse(
modelId="openai.gpt-6-luna",
messages=[{
"role": "user",
"content": [{"text": "Hello"}]
}]
)
print(response["output"]["message"]["content"][0]["text"])Pricing on AWS Bedrock
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.10 |
| Output tokens | $0.50 |
Capabilities
About GPT-6 Luna
OpenAI GPT-6 Luna, announced September 22, 2026 with GPT-6 Sol, is the most efficient GPT-6 tier for focused high-volume tasks below Sol and flagship Astra. First-party OpenAI API id gpt-6-luna. Docs: 1,050,000 context, 128,000 max output, knowledge cutoff May 18, 2026, text and image input to text output, reasoning effort none/low/medium/high/xhigh/max, function calling, structured outputs, prompt caching, Batch and Flex at 50% of Standard, Fast mode at 2x, and computer use and code interpreter via Responses tools.