Llama 3 Models by AI at Meta
Last refreshed 2026-07-09. Next refresh: weekly.
Details
Capabilities
About
Llama 3, developed by Meta AI and released in April 2024, represents a significant advancement in large language models (LLMs). Available in two configurations—8 billion and 70 billion parameters—the models offer both pretrained and instruction-tuned versions, enhancing their adaptability and effectiveness in dialogue scenarios. Llama 3 sets itself apart by being trained on over 15 trillion tokens of publicly available data, a massive expansion over its predecessor, Llama 2, and includes a substantial increase in code data. The models not only excel in performance but also incorporate robust safety features like Llama Guard 2 and Code Shield, underscoring Meta's focus on responsible AI use. Llama 3 models are accessible on platforms such as AWS, Google Cloud, and Hugging Face, with plans for future updates that will expand their capabilities to include multimodal functionalities and multilingual support.
Decision facts
- Best fit
- JSON / Tool usestructured outputscoding
- Capability starting point
- Together AI - Llama 3 8B Lite with 8k context and JSON / Tool use and structured outputs
- Lowest tracked input
- Llama 3 8B Instruct · $0.02/1M · DeepInfra
- Closest related family
- K2 Horizon
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Use when the workload needs 8k context, 8B parameters, and JSON / Tool use.
Use when the workload needs 8k context and 70B parameters.
Use when the workload needs 8k context, 70B parameters, and structured outputs.
Use when the workload needs 8k context, 8B parameters, and structured outputs.
Use when the workload needs 8k context and 70B parameters.
Use when the workload needs 8k context and 8B parameters.
Use when the workload needs 8k context, 8B parameters, and structured outputs.
Use when the workload needs 8k context, 70B parameters, and structured outputs.
Use when the workload needs 8k context, 8B parameters, and structured outputs.
Use when the workload needs 8k context, 70B parameters, and structured outputs.
Use when the workload needs 8k context and 8B parameters.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Together AI - Llama 3 8B Lite | Use when the workload needs 8k context, 8B parameters, and JSON / Tool use. | 2025-07 | 8k context8B parametersJSON / Tool use | Current |
| Llama 3 Taiwan 70B Instruct | Use when the workload needs 8k context and 70B parameters. | 2024-07 | 8k context70B parameters | Current |
| Llama 3 70B Instruct | Use when the workload needs 8k context, 70B parameters, and structured outputs. | 2024-04 | 8k context70B parametersstructured outputs | Current |
| Llama 3 8B Instruct | Use when the workload needs 8k context, 8B parameters, and structured outputs. | 2024-04 | 8k context8B parametersstructured outputs | Current |
| Llama 3 70B | Use when the workload needs 8k context and 70B parameters. | 2024-04 | 8k context70B parameters | Current |
| Llama 3 8B | Use when the workload needs 8k context and 8B parameters. | 2024-04 | 8k context8B parameters | Current |
| Together AI Llama-3-8B-Instruct | Use when the workload needs 8k context, 8B parameters, and structured outputs. | 2024-04 | 8k context8B parametersstructured outputs | Current |
| Together AI Llama-3-70B-Instruct | Use when the workload needs 8k context, 70B parameters, and structured outputs. | 2024-04 | 8k context70B parametersstructured outputs | Current |
| DeepInfra Llama 3 8B Instruct | Use when the workload needs 8k context, 8B parameters, and structured outputs. | 2024-04 | 8k context8B parametersstructured outputs | Current |
| DeepInfra Llama 3 70B Instruct | Use when the workload needs 8k context, 70B parameters, and structured outputs. | 2024-04 | 8k context70B parametersstructured outputs | Current |
| Fireworks Llama-3-8B-Instruct | Use when the workload needs 8k context and 8B parameters. | 2024-04 | 8k context8B parameters | Current |
Release Timeline
3 release groupsSpecifications(11 models)
| Model | Released | Context | Parameters | JSON / Tool use | Structured Outputs |
|---|---|---|---|---|---|
| Together AI - Llama 3 8B Lite | 2025-07 | 8k | 8B | Yes | Yes |
| Llama 3 Taiwan 70B Instruct | 2024-07 | 8k | 70B | No | No |
| Llama 3 70B Instruct | 2024-04 | 8k | 70B | No | Yes |
| Llama 3 8B Instruct | 2024-04 | 8k | 8B | No | Yes |
| Llama 3 70B | 2024-04 | 8k | 70B | No | No |
| Llama 3 8B | 2024-04 | 8k | 8B | No | No |
| Together AI Llama-3-8B-Instruct | 2024-04 | 8k | 8B | No | Yes |
| Together AI Llama-3-70B-Instruct | 2024-04 | 8k | 70B | No | Yes |
| DeepInfra Llama 3 8B Instruct | 2024-04 | 8k | 8B | No | Yes |
| DeepInfra Llama 3 70B Instruct | 2024-04 | 8k | 70B | No | Yes |
| Fireworks Llama-3-8B-Instruct | 2024-04 | 8k | 8B | No | No |
Available From(20 providers)
Pricing
Popular comparisons in this family
- Llama 3 70B Instruct vs Llama 3.1 70B Instruct5K
- Llama 3 70B Instruct vs Qwen2.5-72B-Instruct461
- Llama 3 70B Instruct vs Qwen3.6-27B410
- DeepSeek V4 Flash vs Llama 3 70B Instruct407
- Llama 3 70B Instruct vs Qwen3.6-35B-A3B381
- DeepSeek V4 Pro vs Llama 3 70B Instruct226
- Llama 3 70B Instruct vs Mixtral 8x7B206
- Claude Opus 4.7 vs Llama 3 70B Instruct186
- Claude Sonnet 4.6 vs Llama 3 70B Instruct181
- DeepSeek V4 Flash vs Llama 3 8B Instruct157
Models(11)
Together AI - Llama 3 8B Lite
Llama 3 Taiwan 70B Instruct
Llama 3 70B Instruct
Llama 3 8B Instruct
Llama 3 70B
Llama 3 8B
Together AI Llama-3-8B-Instruct
Together AI Llama-3-70B-Instruct
DeepInfra Llama 3 8B Instruct
DeepInfra Llama 3 70B Instruct
Fireworks Llama-3-8B-Instruct
