Gemma 2B
Gemma 2B is a released coding and classification model with open-weight; evaluate it while provider pricing coverage matures.
Use it for
- Teams evaluating coding and classification
- Workloads that can use a 2k context window
Do not use it for
- Cost-sensitive launches that need sourced token pricing
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
No tracked provider token pricing is available yet.
About
Gemma 2B is Google DeepMind's Gemma model. Weights are openly available for self-hosting and scores 39.8 on GPQA.
Gemma 2B is an open-weight model in the Gemma family. The structured metadata tracks a 2k-token context window. Headline tracked benchmarks include Google-Proof Q&A 39.8, HellaSwag 83.4, and HumanEval 58.9.
Top use-case fit: coding, agents, and build tasks
Coding
1 relevant benchmark in the decision map.
Classification
2 relevant benchmarks in the decision map.
Provider price ladder
No tracked provider token pricing is available for this model yet.
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
Benchmark scores(4)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Google-Proof Q&A | 39.8 | diamondObserved 2026-03-06 | — | research |
| HellaSwag | 83.4 | 10-shotObserved 2026-03-06 | — | Source |
| HumanEval | 58.9 | pass@1Observed 2026-03-06 | — | Source |
| Massive Multitask Language Understanding | 64.2 | 5-shotObserved 2026-03-06 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Frequently asked questions
What is the context window of Gemma 2B?
Gemma 2B has a context window of 2k tokens.
When was Gemma 2B released?
Gemma 2B was released on 2024-02-21.
What benchmarks has Gemma 2B been tested on?
Gemma 2B has been evaluated on 4 benchmarks, including Google-Proof Q&A, HellaSwag, HumanEval, Massive Multitask Language Understanding.
No tracked provider token pricing is available yet.