LLM Reference

NVIDIA Llama 3 ChatQA Models by NVIDIA AI

NVIDIA AILlama 3 CommunityOpen weights
2 models2024Up to 8k ctxFrom $0.37/1M input

Last refreshed 2026-05-19. Next refresh: weekly.

Details

ResearcherNVIDIA AI
Commercial useCommercial use: conditional
Models2
Released2024
Max context8k

About

The NVIDIA Llama 3 ChatQA family of large language models (LLMs) is designed to excel in conversational question answering (QA) and retrieval-augmented generation (RAG). These models are grounded in the Llama 3 base model and leverage an enhanced training methodology from the ChatQA project. A standout feature is their integration of extensive conversational QA data, which enhances their capability to manage tabular data and complex arithmetic calculations. The family offers two primary variants: Llama3-ChatQA-1.5-8B and Llama3-ChatQA-1.5-70B. These variants cater to different performance needs and computational requirements, with the 70B model excelling in reasoning and language understanding. NVIDIA supports these models with comprehensive resources, including benchmark results and detailed documentation, for developers and researchers 14.

Decision facts

Best fit
General model comparison
Capability starting point
NVIDIA Llama 3 ChatQA 8B with 8k context
Lowest tracked input
NVIDIA Llama 3 ChatQA 8B · $0.37/1M · Microsoft Foundry
Closest related family
NVIDIA Nemotron Nano 12B v2 VL

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view1 retired

Use when the workload needs 8k context and 8B parameters.

2024-088k context8B parameters

Release Timeline

1 release group
2024-08
1 current · 1 retired
NVIDIA Llama 3 ChatQA 70B
8k context70B parameters
Archived
NVIDIA Llama 3 ChatQA 8B
8k context8B parameters
Current

Specifications(2 models)

NVIDIA Llama 3 ChatQA model specifications comparison
ModelReleasedContextParameters
NVIDIA Llama 3 ChatQA 8B2024-088k8B

Available From(2 providers)

Pricing

NVIDIA Llama 3 ChatQA model pricing by provider
ModelProviderInput / 1MOutput / 1MType
NVIDIA Llama 3 ChatQA 8BMicrosoft Foundry$0.37$1.1Provisioned