LLM Reference
NVIDIA AI

NVIDIA AI

53 models across 15 families · Latest: Nemotron-Labs TwoTower 30B-A3B Base (2026-06)

Accelerated AI for enterprise solutions

CodingRAGAgentsLong contextVisionClassificationJSON / Tool useHighlight

NVIDIA AI's portfolio covers 50 active models across 14 current families, spanning coding, rag, and agents. Open a model detail page to compare provider routes and sourced benchmarks.

Covers 7 workload areas across 50 active tracked models; last verified 2026-07-02.

Use it for

  • Teams evaluating coding, rag, and agents across this lab's releases
  • Comparing model families before committing to a flagship
  • Migration and pricing follow-ups across 50 tracked models

Do not use it for

  • Choosing a hosting provider without opening a model page for price ladders

Active models

50

Current models from this lab, excluding deprecated ones

Active families

14

Current model families from this lab

Open catalog

45 open

0 open source / 45 open weights

Lowest output price

$0.160 /1M

Cheapest tracked output across active models, per 1M tokens

Latest dated release

2026-06-25

Nemotron-Labs TwoTower 30B-A3B Base

Freshness

2026-07-02

Researched 62d ago

stale

Information

Founded2015
Santa Clara, California, United States

Release cadence

Dated releases
5
Latest date
2026-06-25

Where this lab wins

  • Coding: 4 tracked models with SWE-bench / HumanEval-style scores.
  • RAG: 2 tracked models with ruler / needle retrieval benchmarks.
  • Agentic: 4 tracked models with BFCL, tau-bench, and SWE-bench tool-use coverage.
  • Long-context: 14 tracked models with context-token or InfiniteBench-class signal.

Flagship quality / price signal

Selection source
Best sourced coding Q/$
Coding grade
B
Benchmark / output price
swe-bench-verified 60.47 · $0.450/1M

About

NVIDIA's journey into the realm of artificial intelligence, specifically in the areas of generative AI and large language models (LLMs), is marked by a series of strategic innovations that have equipped the company to lead in the AI landscape. Originally renowned for its high-quality graphics processing units (GPUs) used in gaming and multimedia, NVIDIA shifted gears to address the demands of AI, focusing on quality over sheer volume. This strategic pivot, underscored by the development of pioneering technologies, has attracted tech giants such as Microsoft, Amazon, and Facebook as customers, and has become instrumental in powering large-scale AI applications.

Featured models

ModelReleasedContextInput price ($/1M)Output price ($/1M)LicenseOpenness
Nemotron-Labs TwoTower 30B-A3B Base2026-06-25128k--NVIDIA Open ModelOpen weights
Nemotron 3 Ultra2026-06-041m$0.5$2.2NVIDIA Open ModelOpen weights
Cosmos 3 Nano2026-05-31256k--OpenMDW 1.1Open weights

Model families

Recent releases

  1. Nemotron-Labs TwoTower 30B-A3B Base- 2026-06-25
  2. Nemotron 3 Ultra- 2026-06-04
  3. Cosmos 3 Nano- 2026-05-31
  4. Cosmos 3 Super- 2026-05-31
  5. Cosmos 3 Super Text2Image- 2026-05-31

Top comparisons

Explore related pages

Last reviewed: 2026-07-02. Data sourced from public lab announcements and provider documentation.