Gemma 3 4B Instruct
Released
2025-01-01
Last refreshed
2026-05-19
Status
Researched 134d ago
Open weightsCommercial use: conditionalLong context
Google DeepMind releases · 50 in the last 12 months · this family litChangelog →
Created by
Pricing
Output / 1M
$0.200
Input / 1M
$0.200
Cheapest of 1 route · Fireworks AI
About
Gemma 3 4B Instruct is Google DeepMind's Gemma 3 model. It offers a 128K-token context window with weights openly available for self-hosting.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Fireworks AI | $0.200 | $0.200 | Serverless |
Available via routers & gateways(1)
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Long context
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Created by
Pricing
Output / 1M
$0.200
Input / 1M
$0.200
Cheapest of 1 route · Fireworks AI