GLM-4.7 Flash

Released
2025-01-01
Last refreshed
2026-06-29
Status
Researched 131d ago
Open sourceCommercial use: permittedRAGLong contextClassificationJSON / Tool use
Zhipu AI releases · 15 in the last 12 months · this family litChangelog →
Specifications
Family
GLM-4
Released
2025-01-01
Context
198k
Parameters
30B (3B active)
Architecture
Decoder Only
Specialization
general
Openness
Open source
License
MITOSI-approvedCommercial use: permitted
Weights
Unknown
Code
Unknown
Training
Pretrained
Created by

Chinese AI research lab developing GLM language models.

Beijing, China
Founded 2019
Website
Pricing
Output / 1M
$0.400
Input / 1M
$0.060

Cheapest of 5 routes · Cloudflare Workers AI

About

GLM-4.7 Flash is Tsinghua Knowledge Engineering Group (THUDM)'s GLM-4 model. It offers a 198K-token context window.

Provider price ladder

Compare all 5

Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
Cloudflare Workers AI$0.060$0.400
Serverless
OpenRouter$0.060$0.400
Serverless
Novita AI$0.070$0.400
Serverless
Vercel AI Gateway$0.070$0.400
Serverless

Available via routers & gateways(1)

Capabilities

Structured Outputs

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.