Using Gemma 2 2B Instruct on NVIDIA NIM
Implementation guide · Gemma 2 · Google DeepMind
ServerlessOpen Weights
NVIDIA NIM exposes Gemma 2 2B Instruct through model ID google/gemma-2-2b-it. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-05-19. Next refresh: weekly.
Quick Start
- 1
- 2Use the NVIDIA NIM SDK or REST API to call
google/gemma-2-2b-it— see the documentation for request format. - 3
Code Examples
Pricing on NVIDIA NIM
Capabilities
No model capability flags are currently sourced.
About Gemma 2 2B Instruct
Gemma 2 2B Instruct is Google DeepMind's Gemma 2 model. Weights are openly available for self-hosting and scores 36.8 on GPQA.
Model Specs
Released2024-07-31
Parameters2B
Context8k
ArchitectureDecoder Only