Last refreshed 2026-06-29. Next refresh: weekly.
Why use Hermes 2 Pro Llama 3 8B on Novita AI?
Novita AI offers Hermes 2 Pro Llama 3 8B with pay-as-you-go pricing at $0.14/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.
Compare Hermes 2 Pro Llama 3 8B across 4 providers to find the best fit for your use caseInput / 1M
$0.14
Output / 1M
$0.14
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: hermes-2-pro-llama-3-8bModel ID
hermes-2-pro-llama-3-8bRequest example
Curated snippets for this provider have not been sourced yet.
Gotchas
- Use provider model ID "hermes-2-pro-llama-3-8b", not the LLMReference slug "hermes2-pro-llama3-8b".
Compare Hermes 2 Pro Llama 3 8B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| OctoAI API (Deprecated) | — | — |
| Microsoft Foundry | $0.37 | $1.10 |
| OpenRouter | $0.14 | $0.14 |
| Novita AI | $0.14 | $0.14 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.14 |
| Output tokens | $0.14 |
Capabilities
No model capability flags are currently sourced.
About Hermes 2 Pro Llama 3 8B
8B Hermes model merging Hermes 2 Pro with Llama 3 architecture for superior function calling and structured outputs. Excels in ChatML format multi-turn conversations.
Get Started
Model Specs
Released2023-12-12
Parameters8B
Context8k
ArchitectureDecoder Only
Knowledge cutoff2023-12