NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
With CoreAI, you can start chatting with NVIDIA: Nemotron 3.5 Lightning instantly — no separate subscription needed. CoreAI lets you chat with NVIDIA: Nemotron 3.5 Lightning alongside 300+ other AI models from NVIDIA and other providers like OpenAI, Anthropic, Google, Meta, and more.
Among the 389 chat models on CoreAI, NVIDIA: Nemotron 3.5 Lightning ranks #155 by context window — 262K tokens, roughly 196,608 words or 393 pages of text in a single conversation. On price, it is the #48 cheapest of 127 budget-tier models by input cost. Released 2026-08-11, it is one of the 2 newest additions to the catalog. It supports step-by-step reasoning (thinking mode).
Every model in this table is available on CoreAI — open the web app and run the same prompt through NVIDIA: Nemotron 3.5 Lightning and NVIDIA: Nemotron 3 Super side by side to judge with your own eyes.
A typical question (about 1,000 tokens in, 500 out) runs $0.00022 on NVIDIA: Nemotron 3.5 Lightning's raw per-token rates; a heavy session with a 50K-token document costs about $0.006. On CoreAI you don't juggle these numbers — one subscription covers NVIDIA: Nemotron 3.5 Lightning and 300+ other models under a plan usage limit, with no API keys. See plans.
NVIDIA: Nemotron 3.5 Lightning handles 262K tokens of context — roughly 196,608 words, or about 393 pages of text in one conversation.
NVIDIA: Nemotron 3.5 Lightning's underlying rates are $0.1 per 1M input tokens and $0.25 per 1M output tokens. On CoreAI it is included in one subscription alongside 300+ models, so there is no per-token billing to manage.
No — NVIDIA: Nemotron 3.5 Lightning is text-only. If you need image understanding, its budget-tier peers with vision are listed in the comparison table above.
Yes. NVIDIA: Nemotron 3.5 Lightning can work through problems step by step before answering; on CoreAI you can toggle thinking mode per message.
Open the CoreAI web app or the iOS/Android app, pick NVIDIA: Nemotron 3.5 Lightning in the model selector, and chat — no API key, no per-token billing. One subscription covers NVIDIA: Nemotron 3.5 Lightning and 300+ other models.
The closest alternatives on CoreAI are NVIDIA: Nemotron 3 Super, inclusionAI: Ling 3.0 Tiny (free), inclusionAI: Ling-2.6-flash — similar tier and capability. The fastest way to choose is running one prompt through them side by side in CoreAI's Compare view.
Chat with NVIDIA: Nemotron 3.5 Lightning and 300+ other AI models — all in one app.