LLM Reference
Cloudflare Workers AI

Using GLM-4.7 Flash on Cloudflare Workers AI

Implementation guide · GLM-4 · Zhipu AI

ServerlessOpen Source

Cloudflare Workers AI exposes GLM-4.7 Flash through model ID @cf/zhipu/glm-4.7-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-15. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Cloudflare Workers AI and generate an API key.
  2. 2
    Use the Cloudflare Workers AI SDK or REST API to call @cf/zhipu/glm-4.7-flash — see the documentation for request format.
  3. 3
    You'll be billed $0.06/1M input, $0.40/1M output tokens. See full pricing.

Code Examples

See Cloudflare Workers AI documentation for integration details.

Pricing on Cloudflare Workers AI

TypePrice (per 1M)
Input tokens$0.06
Output tokens$0.40

Capabilities

Structured Outputs

About GLM-4.7 Flash

GLM-4.7 Flash is Tsinghua Knowledge Engineering Group (THUDM)'s GLM-4 model. It offers a 198K-token context window.

Model Specs

Released2025-01-01
Parameters30B (3B active)
Context198k
ArchitectureDecoder Only

Provider

Cloudflare Workers AI
Cloudflare Workers AI

Cloudflare

San Francisco, California, United States