Llama 3.1 Models by AI at Meta
Details
Capabilities
About
The Llama 3.1 family, developed by Meta, features large language models (LLMs) with sizes of 8B, 70B, and 405B parameters 12. The standout 405B model is the largest openly available foundation model, showcasing capabilities that compete with top closed-source models 1. These models boast a 128K token context window, support for eight languages, and enhanced reasoning capabilities, making them versatile for multiple applications. The 405B model is ideal for synthetic data generation and model distillation, highlighting Meta's focus on open-source AI, enabling customization and deployment without needing to share data with Meta 1.
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Use when the workload needs 128k context, 405B parameters, and structured outputs.
Use when the workload needs 128k context, 70B parameters, and structured outputs.
Use when the workload needs 128k context, 8B parameters, and structured outputs.
Use when the workload needs 128k context and 405B parameters.
Use when the workload needs 128k context and 70B parameters.
Use when the workload needs 128k context and 8B parameters.
Use when the workload needs 128k context, 405B parameters, and tool use.
Use when the workload needs 128k context, 70B parameters, and tool use.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Llama 3.1 405B Instruct | Use when the workload needs 128k context, 405B parameters, and structured outputs. | 2024-07 | 128k context405B parametersstructured outputs | Current |
| Llama 3.1 70B Instruct | Use when the workload needs 128k context, 70B parameters, and structured outputs. | 2024-07 | 128k context70B parametersstructured outputs | Current |
| Llama 3.1 8B Instruct | Use when the workload needs 128k context, 8B parameters, and structured outputs. | 2024-07 | 128k context8B parametersstructured outputs | Current |
| Llama 3.1 405B | Use when the workload needs 128k context and 405B parameters. | 2024-07 | 128k context405B parameters | Current |
| Llama 3.1 70B | Use when the workload needs 128k context and 70B parameters. | 2024-07 | 128k context70B parameters | Current |
| Llama 3.1 8B | Use when the workload needs 128k context and 8B parameters. | 2024-07 | 128k context8B parameters | Current |
| Llama 3.1-405B | Use when the workload needs 128k context, 405B parameters, and tool use. | 2024-07 | 128k context405B parameterstool use | Current |
| Llama 3.1-70B | Use when the workload needs 128k context, 70B parameters, and tool use. | 2024-07 | 128k context70B parameterstool use | Current |
Release Timeline
1 release groupSpecifications(8 models)
| Model | Released | Context | Parameters | Fn Calling | Tool Use | Structured Outputs |
|---|---|---|---|---|---|---|
| Llama 3.1 405B Instruct | 2024-07 | 128k | 405B | No | No | Yes |
| Llama 3.1 70B Instruct | 2024-07 | 128k | 70B | No | No | Yes |
| Llama 3.1 8B Instruct | 2024-07 | 128k | 8B | No | No | Yes |
| Llama 3.1 405B | 2024-07 | 128k | 405B | No | No | No |
| Llama 3.1 70B | 2024-07 | 128k | 70B | No | No | No |
| Llama 3.1 8B | 2024-07 | 128k | 8B | No | No | No |
| Llama 3.1-405B | 2024-07 | 128k | 405B | Yes | Yes | Yes |
| Llama 3.1-70B | 2024-07 | 128k | 70B | Yes | Yes | No |
Available From(18 providers)
Pricing
Popular comparisons in this family
- Llama 3 70B Instruct vs Llama 3.1 70B Instruct5K
- Llama 3.1 405B vs Qwen2.5-72B535
- Llama 3.1 70B Instruct vs Qwen3.6-27B325
- Grok-3 vs Llama 3.1 70B Instruct230
- Grok 4 vs Llama 3.1 70B Instruct211
- Llama 3.1 70B Instruct vs Qwen3.6-35B-A3B179
- DeepSeek V3 vs Llama 3.1 405B170
- Llama 3.1 405B vs Qwen2.5-72B169
- Llama 3.1 405B vs Qwen2.5-14B-Instruct168
- Claude Sonnet 4.6 vs Llama 3.1 70B Instruct149
Frequently Asked Questions
- What is Llama 3.1 used for?
- Llama 3.1 is used for agent workflows and tool use, structured outputs, and coding. The family description and listed model capabilities point to those workloads as the best fit.
- How does Llama 3.1 compare to MOSS-Audio?
- Llama 3.1 by AI at Meta is strongest where you need agent workflows and tool use, while MOSS-Audio by MOSI AI is the closest related family to check for multimodal. Llama 3.1 has 8 listed variants and reaches up to 128k context, so compare the specs and pricing tables before choosing a production model.
- Which Llama 3.1 model should I use?
- For the lowest listed input price, start with Llama 3.1 8B Instruct through DeepInfra at $0.02/1M input tokens. For the most capable/latest local choice, evaluate Llama 3.1-405B with 128k context and tool use, function calling, and structured outputs.
Models(8)
Llama 3.1 405B Instruct
Llama 3.1 70B Instruct
Llama 3.1 8B Instruct
Llama 3.1 405B
Llama 3.1 70B
Llama 3.1 8B
Llama 3.1-405B
Llama 3.1-70B
