LLM Reference
Microsoft Foundry

Using NVIDIA Llama 3 ChatQA 8B on Microsoft Foundry

Implementation guide · NVIDIA Llama 3 ChatQA · NVIDIA AI

ProvisionedOpen Weights

Microsoft Foundry exposes NVIDIA Llama 3 ChatQA 8B through model ID nvidia-llama3-chatqa-8b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-05-19. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Microsoft Foundry and generate an API key.
  2. 2
    Use the Microsoft Foundry SDK or REST API to call nvidia-llama3-chatqa-8b — see the documentation for request format.
  3. 3
    You'll be billed $0.37/1M input, $1.10/1M output tokens. See full pricing.

Code Examples

See Microsoft Foundry documentation for integration details.

Pricing on Microsoft Foundry

TypePrice (per 1M)
Input tokens$0.37
Output tokens$1.10

Capabilities

No model capability flags are currently sourced.

About NVIDIA Llama 3 ChatQA 8B

NVIDIA Llama 3 ChatQA 8B is NVIDIA AI's NVIDIA Llama 3 ChatQA model. It was released 2024-08-15.

Model Specs

Released2024-08-15
Parameters8B
Context8k
ArchitectureDecoder Only

Provider

Microsoft Foundry
Microsoft Foundry

Microsoft

Redmond, Washington, United States