Using Llama 3 Swallow 70B Instruct on NVIDIA NIM

Implementation guide · Swallow · Tokyo Institute of Technology

ServerlessOpen Weights

NVIDIA NIM exposes Llama 3 Swallow 70B Instruct through model ID tokyotech-llm/llama-3-swallow-70b-instruct-v0.1. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-30. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at NVIDIA NIM and generate an API key.
  2. 2
    Use the NVIDIA NIM SDK or REST API to call tokyotech-llm/llama-3-swallow-70b-instruct-v0.1 — see the documentation for request format.
  3. 3
    You'll be billed . See full pricing.

Code Examples

See NVIDIA NIM documentation for integration details.

Pricing on NVIDIA NIM

Capabilities

No model capability flags are currently sourced.

About Llama 3 Swallow 70B Instruct

Llama 3 Swallow 70B Instruct is Tokyo Institute of Technology's Swallow model. Its knowledge cutoff is 2023.

Model Specs

Released2024-06-01
Parameters70B
Context4k
ArchitectureDecoder Only
Knowledge cutoff2023

Provider

NVIDIA NIM

NVIDIA

Santa Clara, California, United States