Nemotron 3 Ultra

Released
2026-06-04
Last refreshed
2026-06-15
Status
Researched 111d ago
Open weightsCommercial use: permittedLong context
NVIDIA AI releases · 19 in the last 12 months · this family litChangelog →
Specifications
Released
2026-06-04
Context
1m
Parameters
550B
Architecture
Mixture of Experts
Specialization
general
Openness
Open weights
License
NVIDIA Open ModelCommercial use: permitted
Weights
Available
Code
Unknown
Created by

Accelerated AI for enterprise solutions

Santa Clara, California, United States
Founded 2015
Website
Pricing
Output / 1M
$2.20
Input / 1M
$0.500

Cheapest of 1 route · OpenRouter

About

NVIDIA's open frontier-reasoning model (550B total / 55B active MoE, hybrid Transformer-Mamba). Highest Artificial Analysis Intelligence Index for any US open model (score: 48). 300+ tokens/second. 1M-token context. Announced at Computex 2026. Pricing: ~$0.60/$2.60 per 1M tokens (provider median); free tier on some providers.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
OpenRouter$0.500$2.20
Serverless

Capabilities

Reasoning

Benchmark peer barsfor Long context

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.