GLM-4V 9B
Released
2024-06-05
Last refreshed
2026-06-15
Status
Researched 105d ago
Open sourceCommercial use: permittedMultimodalLong contextVision
GLM-4V 9B is worth evaluating for long context and vision when its provider route and context window match the workload.
Use it for
- Teams evaluating long context and vision
- Workloads that can use a 131k context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Strict JSON or tool-calling flows
Created by
Pricing
Output / 1M
$0.250
Input / 1M
$0.050
Cheapest of 1 route · Replicate API
About
GLM-4V 9B is Tsinghua Knowledge Engineering Group (THUDM)'s GLM-4 model with multimodal text and image input. It offers a 128K-token context window and scores 48.3 on MMMU.
Top use-case fit
Long context
Included by capability and metadata signals in the decision map.
Vision
Q/$ A1 relevant benchmark in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Replicate API | $0.050 | $0.250 | Serverless |
Capabilities
Multimodal
Benchmark peer barsfor Vision
Benchmark scores(1)
Scores are benchmark-specific and are direction-aware: the same numeric gap can mean very different outcomes across suites. Use the leaderboard context and this model's provider route to decide whether the winning margin is meaningful for your workload.
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Massive Multi-discipline Multimodal Understanding | 48.3 | —Observed 2026-04-14 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Compare GLM-4V 9B with other models
Comparison and alternatives
Browse all comparisons →Created by
Pricing
Output / 1M
$0.250
Input / 1M
$0.050
Cheapest of 1 route · Replicate API