GLM-4.7 Flash
Released
2025-01-01
Last refreshed
2026-06-29
Status
Researched 131d ago
Open sourceCommercial use: permittedRAGLong contextClassificationJSON / Tool use
Zhipu AI releases · 15 in the last 12 months · this family litChangelog →
Created by
Pricing
Output / 1M
$0.400
Input / 1M
$0.060
Cheapest of 5 routes · Cloudflare Workers AI
About
GLM-4.7 Flash is Tsinghua Knowledge Engineering Group (THUDM)'s GLM-4 model. It offers a 198K-token context window.
Provider price ladder
Compare all 5Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Cloudflare Workers AI | $0.060 | $0.400 | Serverless |
| OpenRouter | $0.060 | $0.400 | Serverless |
| Novita AI | $0.070 | $0.400 | Serverless |
| Vercel AI Gateway | $0.070 | $0.400 | Serverless |
Available via routers & gateways(1)
Capabilities
Structured Outputs
Benchmark peer barsfor RAG
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Created by
Pricing
Output / 1M
$0.400
Input / 1M
$0.060
Cheapest of 5 routes · Cloudflare Workers AI