Hermes 3 Models by Nous Research
Last refreshed 2026-05-19. Next refresh: weekly.
Details
About
The Hermes 3 family of large language models (LLMs), developed by NousResearch, represents a significant advancement in generalist instruction models 146. Built upon the Llama 3.1 foundation model, Hermes 3 models are available in 8B, 70B, and 405B parameter versions 146. A key design principle is enhanced steerability, achieved through targeted training to precisely follow system and instruction prompts in a neutral and adaptive manner 146. This leads to models that are highly responsive to system prompts, allowing fine-grained control over behavior and persona 8. In addition to instruction following, Hermes 3 features long-term context retention, multi-turn conversation, complex role-playing, internal monologue abilities, and enhanced agentic function-calling 146. These models excel in structured output generation, utilizing XML tags and scratchpads for transparency and accuracy 1. The training data is a carefully curated blend of approximately 390 million tokens, including a significant portion of synthetically generated responses to encourage precise instruction following and nuanced reasoning 13.
Decision facts
- Best fit
- agent workflowsstructured outputschatbot and role-playing use cases
- Capability starting point
- Hermes 3 Llama 3.1 405B with 128k context
- Lowest tracked input
- Hermes 3 Llama 3.1 70B · $0.7/1M · OpenRouter
- Closest related family
- MOSS-Audio
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Use when the workload needs 128k context and 405B parameters.
Use when the workload needs 128k context and 70B parameters.
Use when the workload needs 128k context and 8B parameters.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Hermes 3 Llama 3.1 405B | Use when the workload needs 128k context and 405B parameters. | 2024-11 | 128k context405B parameters | Current |
| Hermes 3 Llama 3.1 70B | Use when the workload needs 128k context and 70B parameters. | 2024-11 | 128k context70B parameters | Current |
| Hermes 3 Llama 3.1 8B | Use when the workload needs 128k context and 8B parameters. | 2024-11 | 128k context8B parameters | Current |
Release Timeline
1 release groupSpecifications(3 models)
| Model | Released | Context | Parameters |
|---|---|---|---|
| Hermes 3 Llama 3.1 405B | 2024-11 | 128k | 405B |
| Hermes 3 Llama 3.1 70B | 2024-11 | 128k | 70B |
| Hermes 3 Llama 3.1 8B | 2024-11 | 128k | 8B |
Available From(1 provider)
Pricing
| Model | Provider | Input / 1M | Output / 1M | Type |
|---|---|---|---|---|
| Hermes 3 Llama 3.1 70B | OpenRouter | $0.7 | $0.7 | Serverless |
| Hermes 3 Llama 3.1 405B | OpenRouter | $1 | $1 | Serverless |

