Last refreshed 2026-07-11. Next refresh: weekly.
Why use Hermes 2 Pro Llama 3 8B on OctoAI API (Deprecated)?
OctoAI API (Deprecated) offers Hermes 2 Pro Llama 3 8B with competitive pricing. OctoAI was a hosted inference platform for running third-party foundation models.
Compare Hermes 2 Pro Llama 3 8B across 4 providers to find the best fit for your use caseInput / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: hermes2-pro-llama3-8bModel ID
hermes2-pro-llama3-8bRequest example
Curated snippets for this provider have not been sourced yet.
Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare Hermes 2 Pro Llama 3 8B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| OctoAI API (Deprecated) | — | — |
| Microsoft Foundry | $0.37 | $1.10 |
| OpenRouter | $0.14 | $0.14 |
| Novita AI | $0.14 | $0.14 |
Capabilities
No model capability flags are currently sourced.
About Hermes 2 Pro Llama 3 8B
8B Hermes model merging Hermes 2 Pro with Llama 3 architecture for superior function calling and structured outputs. Excels in ChatML format multi-turn conversations.
Get Started
Model Specs
Released2023-12-12
Parameters8B
Context8k
ArchitectureDecoder Only
Knowledge cutoff2023-12
Other Providers(3)
Related Models on OctoAI API (Deprecated)
Provider
OctoAI API (Deprecated)OctoAI (acquired by NVIDIA)
All models on OctoAI API (Deprecated) →Provider setup guide →