Fireworks AI exposes GLM-5 through model ID accounts/fireworks/models/glm-5. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-06-30. Next refresh: weekly.
Quick Start
- 1
- 2Use the Fireworks AI SDK or REST API to call
accounts/fireworks/models/glm-5— see the documentation for request format. - 3
Code Examples
pip install openaiFIREWORKS_API_KEYaccounts/fireworks/models/glm-5Fireworks model IDs use "accounts/fireworks/models/{model-name}" format, e.g. "accounts/fireworks/models/llama4-scout-instruct-basic" or "accounts/fireworks/models/deepseek-r1".
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["FIREWORKS_API_KEY"],
base_url="https://api.fireworks.ai/inference/v1"
)
response = client.chat.completions.create(
model="accounts/fireworks/models/glm-5",
messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)Pricing on Fireworks AI
| Type | Price (per 1M) |
|---|---|
| Input tokens | $1.00 |
| Output tokens | $3.20 |
Capabilities
About GLM-5
Flagship open-weight foundation model from Zhipu AI with 744B parameters (40B active per token) in Mixture of Experts architecture. Trained on 28.5T tokens using DeepSeek Sparse Attention on Huawei Ascend hardware. Achieves state-of-the-art performance on coding and agentic benchmarks (SWE-bench Verified: 77.8%). Supports autonomous planning, multi-step tool use, and self-correction.