GLM-5.3-Flash
GLM-5.3-Flash is worth evaluating for coding, rag, and agents when its provider route and context window match the workload.
Use it for
- Teams evaluating coding, rag, and agents
- Workloads that can use a 1m context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Workloads where another current model has stronger sourced task evidence
Cheapest of 1 route · Z.ai · cache read $0.015
About
GLM-5.3-Flash is Z.ai's first native multimodal model in the GLM-5 series. It is a 320B-total / 18B-active Mixture-of-Experts model with a 1M-token context window and 131,072 max output tokens, supporting image, video, and file input plus function calling, tool use, structured outputs, reasoning, and prompt caching. MIT open weights are ungated on Hugging Face at zai-org/GLM-5.3-Flash. Official API model ID: glm-5.3-flash. The same model was previously tested anonymously as ox-alpha; that name is not a separate catalog slug.
Top use-case fit: coding, agents, and build tasks
Coding
Included by capability and metadata signals in the decision map.
RAG
Included by capability and metadata signals in the decision map.
Agents
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| Z.ai | $0.075 | $0.250 | read $0.015 | Serverless |
Capabilities
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Benchmark scores(5)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Terminal-Bench 2.1 | 84.3 | Terminal-Bench 2.1Observed 2026-08-26 | — | Source |
| DeepSWE 1.1 | 63.4 | DeepSWE v1.1Observed 2026-08-26 | — | Source |
| Toolathlon | 78.4 | Toolathlon VerifiedObserved 2026-08-26 | — | Source |
| Humanity's Last Exam | 55.3 | HLE with toolsObserved 2026-08-26 | — | Source |
| Agents' Last Exam | 26.3 | Agents' Last Exam (CLI)Observed 2026-08-26 | — | Source |
Migration checks
No linked migration route is available for this model yet.
API versions
glm-5.3-flashCheapest of 1 route · Z.ai · cache read $0.015