Last refreshed 2026-06-30. Next refresh: weekly.
Why use GLM-5 on Novita AI?
Novita AI offers GLM-5 with pay-as-you-go pricing at $1.00/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.
Compare GLM-5 across 7 providers to find the best fit for your use caseInput / 1M
$1.00
Output / 1M
$3.20
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: glm-5Model ID
glm-5Request example
Curated snippets for this provider have not been sourced yet.
Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare GLM-5 Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Fireworks AI | $1.00 | $3.20 |
| OpenRouter | $0.60 | $2.08 |
| Together AI | $1.00 | $3.20 |
| GCP Vertex AI | $1.00 | $3.20 |
| NVIDIA NIM | — | — |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $1.00 |
| Output tokens | $3.20 |
Capabilities
ReasoningJSON / Tool useStructured OutputsPrompt Caching
About GLM-5
Flagship open-weight foundation model from Zhipu AI with 744B parameters (40B active per token) in Mixture of Experts architecture. Trained on 28.5T tokens using DeepSeek Sparse Attention on Huawei Ascend hardware. Achieves state-of-the-art performance on coding and agentic benchmarks (SWE-bench Verified: 77.8%). Supports autonomous planning, multi-step tool use, and self-correction.
Get Started
Model Specs
Released2026-02-11
Parameters744B total, 40B active
Context200k
ArchitectureMixture of Experts
Knowledge cutoff2025-11