Last refreshed 2026-06-15. Next refresh: weekly.
Why use Nemotron 3 Ultra on OpenRouter?
OpenRouter offers Nemotron 3 Ultra with pay-as-you-go pricing at $0.50/1M input tokens. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.
Input / 1M
$0.50
Output / 1M
$2.20
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: nvidia/nemotron-3-ultra-550b-a55bModel ID
nvidia/nemotron-3-ultra-550b-a55bRequest example
Curated snippets for this provider are not sourced yet. Use OpenRouter documentation with model ID
nvidia/nemotron-3-ultra-550b-a55b.Gotchas
- Use provider model ID "nvidia/nemotron-3-ultra-550b-a55b", not the LLMReference slug "nemotron-3-ultra".
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.50 |
| Output tokens | $2.20 |
Capabilities
Reasoning
About Nemotron 3 Ultra
NVIDIA's open frontier-reasoning model (550B total / 55B active MoE, hybrid Transformer-Mamba). Highest Artificial Analysis Intelligence Index for any US open model (score: 48). 300+ tokens/second. 1M-token context. Announced at Computex 2026. Pricing: ~$0.60/$2.60 per 1M tokens (provider median); free tier on some providers.
Get Started
Model Specs
Released2026-06-04
Parameters550B
Context1m
ArchitectureMixture of Experts