Using Llama 3.1 70B Instruct on NVIDIA NIM

Implementation guide · Llama 3.1 · AI at Meta

ProvisionedOpen Weights

NVIDIA NIM exposes Llama 3.1 70B Instruct through model ID llama3.1-70b-instruct. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-07-09. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at NVIDIA NIM and generate an API key.
  2. 2
    Use the NVIDIA NIM SDK or REST API to call llama3.1-70b-instruct — see the documentation for request format.
  3. 3
    You'll be billed . See full pricing.

Code Examples

See NVIDIA NIM documentation for integration details.

Pricing on NVIDIA NIM

Capabilities

Structured Outputs

About Llama 3.1 70B Instruct

The Llama 3.1 70B Instruct model is a cutting-edge large language model with 70 billion parameters, designed for instruction-following tasks. It features multilingual capabilities, supporting languages like English, German, French, and others. Fine-tuned using supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF), it excels in understanding and responding to user instructions. The model can handle a context length of up to 128k tokens, making it suitable for complex dialogue systems and applications requiring detailed responses.

Model Specs

Released2024-07-23
Parameters70B
Context128k
ArchitectureDecoder Only
Knowledge cutoff2023-12

Provider

NVIDIA NIM

NVIDIA

Santa Clara, California, United States