NV-Embed Models by NVIDIA AI
4 models2024–2025Up to 4k ctx
Last refreshed 2026-05-19. Next refresh: weekly.
About
NV-Embed is NVIDIA's family of specialized embedding and reranking models, including NV-EmbedCode for code retrieval and NV-EmbedQA/RerankQA for question-answering tasks. NVIDIA NIM also hosts BAAI BGE models (such as BGE-M3) as first-class retrieval endpoints in its API catalog.
Decision facts
- Best fit
- embeddingrankingcoding
- Capability starting point
- NV-EmbedCode 7B v1 with 4k context
- Lowest tracked input
- Not tracked
- Closest related family
- NVIDIA Nemotron Nano 12B v2 VL
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
4 in view
NV-EmbedCode 7B v1Current
Use when the workload needs embedding, 4k context, and 7B parameters.
2025-06embedding4k context7B parameters
Llama 3.2 NV EmbedQA 1B v2Current
Use when the workload needs embedding, 4k context, and 1B parameters.
2025-03embedding4k context1B parameters
Llama 3.2 NV RerankQA 1B v2Current
Use when the workload needs ranking, 4k context, and 1B parameters.
2025-03ranking4k context1B parameters
Llama 3.2 NV EmbedQA 1B v1Current
Use when the workload needs embedding, 512 context, and 1B parameters.
2024-10embedding512 context1B parameters
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| NV-EmbedCode 7B v1 | Use when the workload needs embedding, 4k context, and 7B parameters. | 2025-06 | embedding4k context7B parameters | Current |
| Llama 3.2 NV EmbedQA 1B v2 | Use when the workload needs embedding, 4k context, and 1B parameters. | 2025-03 | embedding4k context1B parameters | Current |
| Llama 3.2 NV RerankQA 1B v2 | Use when the workload needs ranking, 4k context, and 1B parameters. | 2025-03 | ranking4k context1B parameters | Current |
| Llama 3.2 NV EmbedQA 1B v1 | Use when the workload needs embedding, 512 context, and 1B parameters. | 2024-10 | embedding512 context1B parameters | Current |
Release Timeline
3 release groups2025-06
1 current
NV-EmbedCode 7B v1
Currentembedding4k context7B parameters
2025-03
2 current
Llama 3.2 NV EmbedQA 1B v2
Currentembedding4k context1B parameters
Llama 3.2 NV RerankQA 1B v2
Currentranking4k context1B parameters
2024-10
1 current
Llama 3.2 NV EmbedQA 1B v1
Currentembedding512 context1B parameters
Specifications(4 models)
| Model | Released | Context | Parameters |
|---|---|---|---|
| NV-EmbedCode 7B v1 | 2025-06 | 4k | 7B |
| Llama 3.2 NV EmbedQA 1B v2 | 2025-03 | 4k | 1B |
| Llama 3.2 NV RerankQA 1B v2 | 2025-03 | 4k | 1B |
| Llama 3.2 NV EmbedQA 1B v1 | 2024-10 | 512 | 1B |
Available From(1 provider)
Popular comparisons in this family
- Llama 3.2 NV RerankQA 1B v2 vs Llama 3.3 Nemotron Super 49B v1246
- Llama 3.2 NV EmbedQA 1B v2 vs MiniCPM 2B57
- Llama 3.1 Nemotron Nano 4B v1.1 vs Llama 3.2 NV RerankQA 1B v254
- Llama 3.1 Nemotron Nano VL 8B v1 vs Llama 3.2 NV RerankQA 1B v244
- Llama 3.2 NV EmbedQA 1B v2 vs Llama 2 7B38
- Llama 3.2 NV EmbedQA 1B v2 vs Mistral 7B v0.130
- Llama 3.1 Nemotron Nano VL 8B v1 vs Llama 3.2 NV EmbedQA 1B v224
- Llama 3.2 NV RerankQA 1B v2 vs Mistral Nemotron23
- Llama 3.2 NV RerankQA 1B v2 vs MiniCPM 2B17
- Llama 3.2 NV RerankQA 1B v2 vs text-davinci15




