OctoML Gemma-2B-it
Released
2024-02-21
Last refreshed
2026-05-19
Status
Researched 120d ago
Open weightsCommercial use: conditional
OctoML Gemma-2B-it is worth evaluating for general LLM work when its provider route and context window match the workload.
Use it for
- Teams evaluating general LLM work
- Workloads that can use a 8k context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
Created by
Pricing
Output / 1M
$0.150
Input / 1M
$0.100
Cheapest of 1 route · OctoML (Deprecated)
About
OctoML Gemma-2B-it is Google DeepMind's Gemma model. It offers an 8K-token context window with weights openly available for self-hosting.
Top use-case fit
No primary decision-task fit is mapped for this model yet.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| OctoML (Deprecated) | $0.100 | $0.150 | Serverless |
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Created by
Pricing
Output / 1M
$0.150
Input / 1M
$0.100
Cheapest of 1 route · OctoML (Deprecated)