GLM-OCR
Last refreshed
2026-07-11
Status
Researched 45d ago
ProprietaryCommercial use: conditionalMultimodalVisionJSON / Tool useVisionExtraction
GLM-OCR is worth evaluating for vision and json / tool use when its provider route and context window match the workload.
Use it for
- Teams evaluating vision and json / tool use
- Buyers comparing 1 tracked provider route
Do not use it for
- Workloads where another current model has stronger sourced task evidence
Specifications
- Family
- GLM-OCR
- Specialization
- vision
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
Created by
Pricing
Output / 1M
-
Input / 1M
-
Cheapest of 1 route · Zhipu AI GLM API
Links
About
GLM-OCR is Zhipu AI's reusable vision model for optical character recognition, document understanding, layout parsing, and structured document extraction through an API.
Top use-case fit
Vision
Included by capability and metadata signals in the decision map.
JSON / Tool use
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Zhipu AI GLM API | - | - | ServerlessPartial |
Capabilities
VisionMultimodalStructured Outputs
Benchmark peer barsfor Vision
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Created by
Pricing
Output / 1M
-
Input / 1M
-
Cheapest of 1 route · Zhipu AI GLM API
Links