How does Nemotron 3 Ultra rank among AI models?
Nemotron 3 Ultra has no public ModelCap Index position as of 22 September 2026: the snapshot lacks enough matched benchmark evidence, or the identity match is unresolved. Its catalogue facts (price, context, providers) are shown without a rank.
How much does Nemotron 3 Ultra cost per 1M tokens?
Nemotron 3 Ultra is listed at $0.60 per 1M input tokens and $2.40 per 1M output tokens as of 22 September 2026; the lowest output price among 3 providers is $2.20 on DeepInfra. Prices come from the OpenRouter catalogue and re-observe on every refresh; at least one provider lists a free tier. For 1,000 requests using 2,000 input and 500 output tokens each, the listed rates imply $2.40 total (2M input + 0.5M output). This excludes caching, batch discounts, prompt-length tiers, tool charges and retries; verify the chosen endpoint's rate and limits before budgeting.
What is the context window of Nemotron 3 Ultra?
Nemotron 3 Ultra has a published context window of 262K tokens (262,144) and a maximum output of 183K tokens. Individual providers can serve less than the published maximum; the providers table lists each endpoint's own limit.
Is Nemotron 3 Ultra open-weight?
Restricted license (openmdw-1.1). Downloadable weights use a model-specific or use-restricted license. Review the linked license before use, redistribution, or commercial deployment. The repository is nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 on Hugging Face.
Which API providers serve Nemotron 3 Ultra?
3 providers serve Nemotron 3 Ultra through OpenRouter as of 22 September 2026: DeepInfra, BaseTen and Venice. Each provider's price, context limit, quantization and measured uptime are in the providers table above.