LLM Reference

Granite Guardian 4.1 8B

Released
2026-04-29
Last refreshed
2026-09-03
Status
Researched 135d ago
Open sourceCommercial use: permittedClassification

Granite Guardian 4.1 8B is a released classification model with open-source; evaluate it while provider pricing coverage matures.

Use it for

  • Teams evaluating classification
  • Workloads that can use a 8k context window

Do not use it for

  • Cost-sensitive launches that need sourced token pricing
  • Vision or document-understanding workloads
  • Strict JSON or tool-calling flows
Specifications
Released
2026-04-29
Context
8k
Parameters
8B
Specialization
safety
Openness
Open source
License
Apache 2.0OSI-approvedCommercial use: permitted
Weights
Available
Code
Unknown
Created by

Creating reliable and adaptable AI solutions

Armonk, New York, United States
Founded 1945
Website
Pricing

No tracked provider token pricing is available yet.

About

IBM Granite Guardian 4.1 8B is a safety and risk-detection model fine-tuned from Granite 4.1 8B. Features hybrid thinking mode (detailed reasoning via <think> tags or fast yes/no). Detects: harm, social bias, jailbreaking, violence, profanity, sexual content, unethical behavior, RAG hallucination (groundedness, context relevance, answer relevance), and function-calling hallucination in agentic workflows. Supports custom judging criteria (BYOC). Acts as reward model for best-of-N selection. Benchmarks: OOD Safety F1 0.79, RAG hallucination avg BAcc 0.76, BFCL function calling BAcc 0.79. Apache 2.0.

Top use-case fit

Classification

Included by capability and metadata signals in the decision map.

Provider price ladder

No tracked provider token pricing is available for this model yet.

Capabilities

No model capability flags are currently sourced.

Benchmark peer barsfor Classification

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.