Gemma 3 4B Instruct

Released
2025-01-01
Last refreshed
2026-05-19
Status
Researched 134d ago
Open weightsCommercial use: conditionalLong context
Google DeepMind releases · 50 in the last 12 months · this family litChangelog →
Specifications
Family
Gemma 3
Released
2025-01-01
Context
128k
Parameters
4B
Architecture
Decoder Only
Knowledge cutoff
2024-08
Specialization
general
Openness
Open weights
License
GemmaCommercial use: conditional
Weights
Unknown
Code
Unknown
Training
Pretrained
Created by

Pioneering artificial intelligence research.

London, United Kingdom
Founded 2014
Website
Pricing
Output / 1M
$0.200
Input / 1M
$0.200

Cheapest of 1 route · Fireworks AI

About

Gemma 3 4B Instruct is Google DeepMind's Gemma 3 model. It offers a 128K-token context window with weights openly available for self-hosting.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
Fireworks AI$0.200$0.200
Serverless

Available via routers & gateways(1)

Capabilities

No model capability flags are currently sourced.

Benchmark peer barsfor Long context

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.