Llama 3.1 8B Instruct on IBM watsonx

Llama 3.1 · AI at Meta

ServerlessOpen Weights

Last refreshed 2026-07-09. Next refresh: weekly.

Why use Llama 3.1 8B Instruct on IBM watsonx?

IBM watsonx offers Llama 3.1 8B Instruct with pay-as-you-go pricing at $0.15/1M input tokens. IBM offers a comprehensive AI platform that includes IBM watsonx, a next-generation enterprise AI and data platform.

Compare Llama 3.1 8B Instruct across 17 providers to find the best fit for your use case
Input / 1M
$0.15
Output / 1M
$0.50
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: llama3.1-8b-instruct
Model ID
llama3.1-8b-instruct

Request example

Curated snippets for this provider are not sourced yet. Use IBM watsonx documentation with model ID llama3.1-8b-instruct.

Gotchas

No curated gotchas have been sourced for this exact provider/model route yet.

Compare Llama 3.1 8B Instruct Across Providers

ProviderInput (per 1M)Output (per 1M)
Cloudflare Workers AI——
OctoAI API (Deprecated)——
Together AI$0.18$0.18
Fireworks AI$0.20$0.20
NVIDIA NIM——
View all 17 providers →

Pricing

TypePrice (per 1M)
Input tokens$0.15
Output tokens$0.50

Capabilities

Structured Outputs

About Llama 3.1 8B Instruct

The Llama 3.1 8B Instruct model, released on July 23, 2024, is a multilingual large language model with 8 billion parameters, optimized for instruction-following tasks. It features an enhanced transformer architecture, supporting languages like English, German, French, and others. The model excels in dialogue applications, having been fine-tuned using supervised fine-tuning and reinforcement learning with human feedback. Trained on approximately 15 trillion tokens with a December 2023 data cutoff, it outperforms many existing open-source and closed chat models in various benchmarks.

Get Started