LLM Reference
OpenRouter

GLM-5.2 on OpenRouter

GLM-5 · Zhipu AI

ServerlessOpen Source

Last refreshed 2026-06-24. Next refresh: weekly.

Why use GLM-5.2 on OpenRouter?

OpenRouter offers GLM-5.2 with pay-as-you-go pricing at $1.40/1M input tokens. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.

Input / 1M
$1.40
Output / 1M
$4.40
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: z-ai/glm-5.2
Model ID
z-ai/glm-5.2

Request example

Curated snippets for this provider are not sourced yet. Use OpenRouter documentation with model ID z-ai/glm-5.2.

Gotchas

  • Use provider model ID "z-ai/glm-5.2", not the LLMReference slug "glm-5.2".

Pricing

TypePrice (per 1M)
Input tokens$1.40
Output tokens$4.40

Capabilities

ReasoningJSON / Tool useStructured OutputsCode Execution

About GLM-5.2

GLM-5.2 is Z.ai's coding-first successor to GLM-5.1 in the GLM-5 family, released June 13 2026. 753B parameters (40B active) in IndexShare MoE architecture; the IndexShare innovation reuses the same attention indexer across every four sparse layers, cutting per-token FLOPs by 2.9x at 1M context length. Trained on 28.5T tokens. Supports a 1M-token context window via the glm-5.2[1m] model ID, with 131,072-token maximum output and High/Max thinking-effort levels designed for extended agentic coding sessions. MIT license; open weights available on Hugging Face (zai-org/GLM-5.2 and zai-org/GLM-5.2-FP8).

Get Started

Model Specs

Released2026-06-13
Parameters753B total, 40B active
Context1m
ArchitectureMixture of Experts

Related Models on OpenRouter