NVIDIA AI
53 models across 15 families · Latest: Nemotron-Labs TwoTower 30B-A3B Base (2026-06)
Accelerated AI for enterprise solutions
NVIDIA AI's portfolio covers 50 active models across 14 current families, spanning coding, rag, and agents. Open a model detail page to compare provider routes and sourced benchmarks.
Covers 7 workload areas across 50 active tracked models; last verified 2026-07-02.
Use it for
- Teams evaluating coding, rag, and agents across this lab's releases
- Comparing model families before committing to a flagship
- Migration and pricing follow-ups across 50 tracked models
Do not use it for
- Choosing a hosting provider without opening a model page for price ladders
Active models
50
Current models from this lab, excluding deprecated ones
Active families
14
Current model families from this lab
Open catalog
45 open
0 open source / 45 open weights
Lowest output price
$0.160 /1M
Cheapest tracked output across active models, per 1M tokens
Latest dated release
2026-06-25
Nemotron-Labs TwoTower 30B-A3B Base
Freshness
2026-07-02
Researched 62d ago
Information
Release cadence
- Dated releases
- 5
- Latest model
- Nemotron-Labs TwoTower 30B-A3B Base
- Latest date
- 2026-06-25
Where this lab wins
- Coding: 4 tracked models with SWE-bench / HumanEval-style scores.
- RAG: 2 tracked models with ruler / needle retrieval benchmarks.
- Agentic: 4 tracked models with BFCL, tau-bench, and SWE-bench tool-use coverage.
- Long-context: 14 tracked models with context-token or InfiniteBench-class signal.
Flagship quality / price signal
- Flagship
- Nemotron 3 Super-120B-A12B
- Selection source
- Best sourced coding Q/$
- Coding grade
- B
- Benchmark / output price
- swe-bench-verified 60.47 · $0.450/1M
About
NVIDIA's journey into the realm of artificial intelligence, specifically in the areas of generative AI and large language models (LLMs), is marked by a series of strategic innovations that have equipped the company to lead in the AI landscape. Originally renowned for its high-quality graphics processing units (GPUs) used in gaming and multimedia, NVIDIA shifted gears to address the demands of AI, focusing on quality over sheer volume. This strategic pivot, underscored by the development of pioneering technologies, has attracted tech giants such as Microsoft, Amazon, and Facebook as customers, and has become instrumental in powering large-scale AI applications.
Featured models
| Model | Released | Context | Input price ($/1M) | Output price ($/1M) | License | Openness |
|---|---|---|---|---|---|---|
| Nemotron-Labs TwoTower 30B-A3B Base | 2026-06-25 | 128k | - | - | NVIDIA Open Model | Open weights |
| Nemotron 3 Ultra | 2026-06-04 | 1m | $0.5 | $2.2 | NVIDIA Open Model | Open weights |
| Cosmos 3 Nano | 2026-05-31 | 256k | - | - | OpenMDW 1.1 | Open weights |
Model families
Recent releases
- Nemotron-Labs TwoTower 30B-A3B Base- 2026-06-25
- Nemotron 3 Ultra- 2026-06-04
- Cosmos 3 Nano- 2026-05-31
- Cosmos 3 Super- 2026-05-31
- Cosmos 3 Super Text2Image- 2026-05-31
Top comparisons
- Llama 3.2 NV RerankQA 1B v2 vs Llama 3.3 Nemotron Super 49B v1246
- Hunyuan Hy3 Preview vs Llama 3.3 Nemotron Super 49B v1237
- Gemma 3 vs Nemotron 3 Nano Omni221
- Hunyuan Hy3 Preview vs Nemotron Mini 4B Instruct202
- Llama 3.1 NemoGuard 8B Content Safety vs text-curie68
- Llama 3.1 NemoGuard 8B Content Safety vs ShieldGemma 9B65
- Llama 3.2 NV EmbedQA 1B v2 vs MiniCPM 2B57
- Llama 3.3 Nemotron Super 49B v1 vs Nemotron Mini 4B Instruct55
Explore related pages
Last reviewed: 2026-07-02. Data sourced from public lab announcements and provider documentation.










