Using Llama 3.3 70B Instruct (free) on NVIDIA NIM

Implementation guide · Llama 3.3 · AI at Meta

ServerlessOpen Weights

NVIDIA NIM exposes Llama 3.3 70B Instruct (free) through model ID meta/llama-3.3-70b-instruct. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-07-09. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at NVIDIA NIM and generate an API key.
  2. 2
    Use the NVIDIA NIM SDK or REST API to call meta/llama-3.3-70b-instruct — see the documentation for request format.
  3. 3
    You'll be billed . See full pricing.

Code Examples

See NVIDIA NIM documentation for integration details.

Pricing on NVIDIA NIM

Capabilities

Structured Outputs

About Llama 3.3 70B Instruct (free)

Meta: Llama 3.3 70B Instruct (free) available via OpenRouter. Pricing: $null/1M input, $null/1M output.

Model Specs

Released2024-12-06
Parameters70B
Context66k
ArchitectureDecoder Only
Knowledge cutoff2023-12

Provider

NVIDIA NIM

NVIDIA

Santa Clara, California, United States