NVIDIA Nemotron 3 Ultra now available in Command Code! Strongest US open model yet! 🍀
- 1M context
- 5x faster inference
- 550B MoE frontier-intelligence open model
DEAL 2.3x usage 🎟️. $1 Go plan gets you ~$23 usage on Nemotron. Woah, it's fast x taste compliance is great!
Nemotron 3 Ultra at a glance
Open reasoning model for long-horizon autonomous agents.
| Spec | Nemotron 3 Ultra |
|---|---|
| Model ID | nvidia/nemotron-3-ultra-550b-a55b |
| Context window | 1M tokens |
| Released | June 4, 2026 |
| Intelligence Index | 22.9 |
| Coding Index | 49.3 |
| Input | $0.60 per 1M tokens |
| Output | $2.40 per 1M tokens |
| Cache read | $0.12 per 1M tokens |
| Effective input in an agent loop | $0.26 per 1M tokens |
| Plans | Go and above |
Specs, pricing and plans are from the Nemotron 3 Ultra model page as of September 28, 2026, including any deal running that day, so they can differ from the launch pricing above. Intelligence and Coding Index scores are from Artificial Analysis.
Use Nemotron 3 Ultra
1cmd --model nvidia/nemotron-3-ultra-550b-a55bOr type /model in a session and pick it. You can switch mid-session without losing context.
What real work costs on Nemotron 3 Ultra
| Task | Tokens (in · out) | Cost |
|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0037 |
| Review a 500-line PR | 60K · 4K | $0.03 |
| Fix a bug (agent loop) | 180K · 12K | $0.07 |
| Refactor a module | 320K · 20K | $0.13 |
| Full-repo agent run | 900K · 45K | $0.31 |
Try Command Code
The best coding agent for open models, in your terminal or on your desktop.
1npm i -g command-codeDesktop App · Docs · X · Discord
