Granite 4.1 8B
Granite 4.1 8B is worth evaluating for coding, agents, and long context when its provider route and context window match the workload.
Use it for
- Teams evaluating coding, agents, and long context
- Workloads that can use a 131k context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- Granite 4.1
- Released
- 2026-04-29
- Context
- 131k
- Parameters
- 8B
- Architecture
- Decoder Only
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
Cheapest of 1 route · OpenRouter
About
IBM Granite 4.1 8B is a dense decoder-only transformer instruct model with 40 layers, 4096 embedding size, GQA (32 attention heads, 8 KV heads). Supports multilingual dialog (12 languages), code with FIM, tool-calling/function-calling, RAG, and summarization. Trained on NVIDIA GB200 NVL72 cluster. Apache 2.0. Benchmarks: MMLU 73.84, HumanEval 85.37, GSM8K 92.49, BFCL v3 68.27.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ A2 relevant benchmarks in the decision map.
Agents
Q/$ A1 relevant benchmark in the decision map.
Long context
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| OpenRouter | $0.050 | $0.100 | Serverless |
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
Benchmark scores(7)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Berkeley Function Calling Leaderboard v3 | 68.3 | BFCL v3 (accuracy)Observed 2026-06-07 | — | Source |
| BigCodeBench | 35.0 | BigCodeBench (pass@1)Observed 2026-06-07 | — | Source |
| Google-Proof Q&A | 42.0 | 0-shot CoT, instruct model (accuracy)Observed 2026-06-07 | — | Source |
| Grade School Math 8K | 92.5 | 8-shot, instruct model (accuracy)Observed 2026-06-07 | — | Source |
| HumanEval | 87.2 | pass@1, instruct model (pass@1)Observed 2026-06-07 | — | Source |
| Massive Multitask Language Understanding | 73.8 | 5-shot, instruct model (accuracy)Observed 2026-06-07 | — | Source |
| MMLU PRO | 56.0 | 5-shot CoT, instruct model (accuracy)Observed 2026-06-07 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(1)
Cheapest of 1 route · OpenRouter