Using Llama 3.1 Nemotron Nano VL 8B v1 on NVIDIA NIM

Implementation guide · Nemotron Nano 2 · NVIDIA AI

ServerlessOpen Weights

NVIDIA NIM exposes Llama 3.1 Nemotron Nano VL 8B v1 through model ID nvidia/llama-3.1-nemotron-nano-vl-8b-v1. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-09-22. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at NVIDIA NIM and generate an API key.
  2. 2
    Use the NVIDIA NIM SDK or REST API to call nvidia/llama-3.1-nemotron-nano-vl-8b-v1 — see the documentation for request format.
  3. 3
    You'll be billed . See full pricing.

Code Examples

See NVIDIA NIM documentation for integration details.

Pricing on NVIDIA NIM

Capabilities

VisionMultimodal

About Llama 3.1 Nemotron Nano VL 8B v1

Vision-language variant of NVIDIA Nemotron Nano 8B with multimodal capabilities.

Model Specs

Released2025-03-01
Parameters8B
Context4k
ArchitectureDecoder Only

Provider

NVIDIA NIM

NVIDIA

Santa Clara, California, United States