LLM Reference
Microsoft Foundry

Llama 3.1 8B Instruct on Microsoft Foundry

Llama 3.1 · AI at Meta

ProvisionedOpen Weights

Last refreshed 2026-07-09. Next refresh: weekly.

Why use Llama 3.1 8B Instruct on Microsoft Foundry?

Microsoft Foundry offers Llama 3.1 8B Instruct with pay-as-you-go pricing at $0.30/1M input tokens. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.

Compare Llama 3.1 8B Instruct across 16 providers to find the best fit for your use case
Input / 1M
$0.30
Output / 1M
$0.61
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: llama3.1-8b-instruct
Model ID
llama3.1-8b-instruct

Request example

Curated snippets for this provider are not sourced yet. Use Microsoft Foundry documentation with model ID llama3.1-8b-instruct.

Gotchas

No curated gotchas have been sourced for this exact provider/model route yet.

Compare Llama 3.1 8B Instruct Across Providers

ProviderInput (per 1M)Output (per 1M)
Cloudflare Workers AI
OctoAI API (Deprecated)
Together AI$0.18$0.18
Fireworks AI$0.20$0.20
NVIDIA NIM
View all 16 providers →

Pricing

TypePrice (per 1M)
Input tokens$0.30
Output tokens$0.61

Capabilities

Structured Outputs

About Llama 3.1 8B Instruct

The Llama 3.1 8B Instruct model, released on July 23, 2024, is a multilingual large language model with 8 billion parameters, optimized for instruction-following tasks. It features an enhanced transformer architecture, supporting languages like English, German, French, and others. The model excels in dialogue applications, having been fine-tuned using supervised fine-tuning and reinforcement learning with human feedback. Trained on approximately 15 trillion tokens with a December 2023 data cutoff, it outperforms many existing open-source and closed chat models in various benchmarks.

Get Started