Using Gemma 2 9B Instruct on NVIDIA NIM
Implementation guide · Gemma 2 · Google DeepMind
NVIDIA NIM exposes Gemma 2 9B Instruct through model ID gemma-2-9b-it. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-05-01. Next refresh: weekly.
Quick Start
- 1
- 2Use the NVIDIA NIM SDK or REST API to call
gemma-2-9b-it— see the documentation for request format. - 3
Code Examples
Pricing on NVIDIA NIM
Capabilities
About Gemma 2 9B Instruct
Gemma 2 9B Instruct, developed by Google, is a state-of-the-art large language model based on the advanced Gemini framework. It is a decoder-only transformer model with 9 billion parameters, offering a balance between size and performance. The model is trained on an expansive dataset comprising 8 trillion tokens, including web documents, code, and mathematical text, a notable 30% increase from its predecessor, Gemma 1.1. This allows it to adeptly handle diverse tasks such as question answering, creative writing, coding, and mathematical problem-solving.