LLM Reference
Microsoft Foundry

Using Hermes 2 Pro Llama 3 8B on Microsoft Foundry

Implementation guide · Hermes 2 · Nous Research

ProvisionedOpen Source

Microsoft Foundry exposes Hermes 2 Pro Llama 3 8B through model ID hermes2-pro-llama3-8b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-05-19. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Microsoft Foundry and generate an API key.
  2. 2
    Use the Microsoft Foundry SDK or REST API to call hermes2-pro-llama3-8b — see the documentation for request format.
  3. 3
    You'll be billed $0.37/1M input, $1.10/1M output tokens. See full pricing.

Code Examples

See Microsoft Foundry documentation for integration details.

Pricing on Microsoft Foundry

TypePrice (per 1M)
Input tokens$0.37
Output tokens$1.10

Capabilities

No model capability flags are currently sourced.

About Hermes 2 Pro Llama 3 8B

8B Hermes model merging Hermes 2 Pro with Llama 3 architecture for superior function calling and structured outputs. Excels in ChatML format multi-turn conversations.

Model Specs

Released2023-12-12
Parameters8B
Context8k
ArchitectureDecoder Only
Knowledge cutoff2023-12

Provider

Microsoft Foundry
Microsoft Foundry

Microsoft

Redmond, Washington, United States