NVIDIA

NVIDIA: Nemotron 3 Ultra

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

Context Window
256K tokens
Parameters
55B
Input Price
$0.63/1M
Output Price
$3.13/1M
Price Tier
standard
Provider
NVIDIA

How to Use NVIDIA: Nemotron 3 Ultra

With CoreAI, you can start chatting with NVIDIA: Nemotron 3 Ultra instantly — no separate subscription needed. CoreAI lets you chat with NVIDIA: Nemotron 3 Ultra alongside 300+ other AI models from NVIDIA and other providers like OpenAI, Anthropic, Google, Meta, and more.

  1. Download the CoreAI app for iOS, Android, or use the Web App
  2. Select NVIDIA: Nemotron 3 Ultra from the model selector
  3. Start chatting, comparing, or creating with AI

NVIDIA: Nemotron 3 Ultra at a Glance

NVIDIA: Nemotron 3 Ultra holds the #244 spot for context size out of 412 chat models in the CoreAI catalog — 256K tokens, roughly 192,000 words or 384 pages of text in a single conversation. On price, it is the #87 cheapest of 198 standard-tier models by input cost. It joined the catalog on 2026-06-04. It supports step-by-step reasoning (thinking mode).

NVIDIA: Nemotron 3 Ultra Compared to Its Closest Alternatives

SpecNVIDIA: Nemotron 3 UltraNVIDIA: Nemotron 3.5 Content SafetyStepFun: Step 3.7 FlashIBM: Granite 4.0 Micro
ProviderNVIDIANVIDIAStepfunIbm-granite
Input $/1M tokens$0.625$0.2$0.2$0.017
Output $/1M tokens$3.13$0.2$1.15$0.112
Context window256K131K256K131K
Reasoning modeYesYesYesNo
Image inputNoYesYesNo
Price tierstandardbudgetstandardbudget

Every model in this table is available on CoreAI — open the web app and run the same prompt through NVIDIA: Nemotron 3 Ultra and NVIDIA: Nemotron 3.5 Content Safety side by side to judge with your own eyes.

What NVIDIA: Nemotron 3 Ultra Actually Costs in Practice

A typical question (about 1,000 tokens in, 500 out) runs $0.00219 on NVIDIA: Nemotron 3 Ultra's raw per-token rates; a heavy session with a 50K-token document costs about $0.047. On CoreAI you don't juggle these numbers — one subscription covers NVIDIA: Nemotron 3 Ultra and 300+ other models under a plan usage limit, with no API keys. See plans.

More NVIDIA Models

NVIDIA

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throu
262K budget
NVIDIA

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throu
1000K budget
NVIDIA

NVIDIA: Nemotron 3.5 Content Safety

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates b
131K budget
NVIDIA

NVIDIA: Nemotron 3.5 Content Safety (free)

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates b
128K budget
NVIDIA

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built o
1000K budget
NVIDIA

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems
256K budget

Frequently Asked Questions about NVIDIA: Nemotron 3 Ultra

What is NVIDIA: Nemotron 3 Ultra's context window?

NVIDIA: Nemotron 3 Ultra handles 256K tokens of context — roughly 192,000 words, or about 384 pages of text in one conversation.

How much does NVIDIA: Nemotron 3 Ultra cost?

NVIDIA: Nemotron 3 Ultra's underlying rates are $0.625 per 1M input tokens and $3.13 per 1M output tokens. On CoreAI it is included in one subscription alongside 300+ models, so there is no per-token billing to manage.

Can NVIDIA: Nemotron 3 Ultra process images?

No — NVIDIA: Nemotron 3 Ultra is text-only. If you need image understanding, its standard-tier peers with vision are listed in the comparison table above.

Does NVIDIA: Nemotron 3 Ultra support reasoning or thinking mode?

Yes. NVIDIA: Nemotron 3 Ultra can work through problems step by step before answering; on CoreAI you can toggle thinking mode per message.

How can I use NVIDIA: Nemotron 3 Ultra without an API key?

Open the CoreAI web app or the iOS/Android app, pick NVIDIA: Nemotron 3 Ultra in the model selector, and chat — no API key, no per-token billing. One subscription covers NVIDIA: Nemotron 3 Ultra and 300+ other models.

What are the best alternatives to NVIDIA: Nemotron 3 Ultra?

The closest alternatives on CoreAI are NVIDIA: Nemotron 3.5 Content Safety, StepFun: Step 3.7 Flash, IBM: Granite 4.0 Micro — similar capability. The fastest way to choose is running one prompt through them side by side in CoreAI's Compare view.

Read More about NVIDIA: Nemotron 3 Ultra

Try NVIDIA: Nemotron 3 Ultra Now

Chat with NVIDIA: Nemotron 3 Ultra and 300+ other AI models — all in one app.

Use Web App Download App →