Nemotron 3 Ultra
Released
2026-06-04
Last refreshed
2026-06-15
Status
Researched 111d ago
Open weightsCommercial use: permittedLong context
NVIDIA AI releases · 19 in the last 12 months · this family litChangelog →
Specifications
- Family
- Nemotron 3
- Released
- 2026-06-04
- Context
- 1m
- Parameters
- 550B
- Architecture
- Mixture of Experts
- Specialization
- general
- Openness
- Open weights
- License
- NVIDIA Open ModelCommercial use: permitted
- Weights
- Available
- Code
- Unknown
Created by
Pricing
Output / 1M
$2.20
Input / 1M
$0.500
Cheapest of 1 route · OpenRouter
Links
About
NVIDIA's open frontier-reasoning model (550B total / 55B active MoE, hybrid Transformer-Mamba). Highest Artificial Analysis Intelligence Index for any US open model (score: 48). 300+ tokens/second. 1M-token context. Announced at Computex 2026. Pricing: ~$0.60/$2.60 per 1M tokens (provider median); free tier on some providers.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| OpenRouter | $0.500 | $2.20 | Serverless |
Capabilities
Reasoning
Benchmark peer barsfor Long context
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Created by
Pricing
Output / 1M
$2.20
Input / 1M
$0.500
Cheapest of 1 route · OpenRouter
Links