LLM Reference
Microsoft Foundry

Llama 3.1 70B Instruct on Microsoft Foundry

Llama 3.1 · AI at Meta

ProvisionedOpen Weights

Last refreshed 2026-07-09. Next refresh: weekly.

Why use Llama 3.1 70B Instruct on Microsoft Foundry?

Microsoft Foundry offers Llama 3.1 70B Instruct with pay-as-you-go pricing at $2.68/1M input tokens. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.

Compare Llama 3.1 70B Instruct across 13 providers to find the best fit for your use case
Input / 1M
$2.68
Output / 1M
$3.54
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: llama3.1-70b-instruct
Model ID
llama3.1-70b-instruct

Request example

Curated snippets for this provider are not sourced yet. Use Microsoft Foundry documentation with model ID llama3.1-70b-instruct.

Gotchas

No curated gotchas have been sourced for this exact provider/model route yet.

Compare Llama 3.1 70B Instruct Across Providers

ProviderInput (per 1M)Output (per 1M)
Cloudflare Workers AI
OctoAI API (Deprecated)
Together AI$0.88$0.88
Fireworks AI$0.90$0.90
NVIDIA NIM
View all 13 providers →

Pricing

TypePrice (per 1M)
Input tokens$2.68
Output tokens$3.54

Capabilities

Structured Outputs

About Llama 3.1 70B Instruct

The Llama 3.1 70B Instruct model is a cutting-edge large language model with 70 billion parameters, designed for instruction-following tasks. It features multilingual capabilities, supporting languages like English, German, French, and others. Fine-tuned using supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF), it excels in understanding and responding to user instructions. The model can handle a context length of up to 128k tokens, making it suitable for complex dialogue systems and applications requiring detailed responses.

Get Started