NewNemotron 3.5 Lightning

NVIDIA API model specs with normalized input price, output price, context window, sample workload cost, and related comparisons.

Page updated:  Data confirmed:  Prices normalized to USD per 1M tokens Sample workload: 1M input + 500K output

Budget brief

Nemotron 3.5 Lightning is estimated at $0.18 for the standard workload.

Nemotron 3.5 Lightning is best suited for cost-sensitive production traffic. A 1M input token plus 500K output token workload is estimated at $0.18.

Input price$0.08per 1M tokens
Output price$0.2per 1M tokens
Context window262.14Kreported maximum
Releasecatalog timestamp

Use this as a first-pass planning estimate, then verify provider limits, routing, discounts, and availability before production deployment.

Model Specs
ProviderNVIDIA
Model IDnvidia/nemotron-3.5-lightning
Prompt Price
per 1M tokens
$0.08
Completion Price
per 1M tokens
$0.2
Sample Workload Cost
1M input + 500K output
$0.18
Context Window262.14K
Release Date

Estimate your workload cost

Estimate this model for your workload

Prices are normalized to USD per 1M tokens.
Nemotron 3.5 Lightning Calculating… Estimated monthly API cost
Unit prices $0.08 input / $0.2 output Per 1M tokens

This estimate uses normalized public API pricing per 1M tokens. It is a planning aid, not a billing quote. Verify provider pricing, limits, and terms before production use.

Alternative path

Alternative Shortlist

Open model finder

Use these rows to build a test shortlist around Nemotron 3.5 Lightning. Lower cost, same-provider fit, and larger context are separate decisions, so each group is ranked by a different signal.

Lower-cost alternatives

Cross-provider candidates with a lower standard 1M input plus 500K output estimate.

ModelProviderSample CostContextWhy it is hereNext Step
New🔥Ox Alphastealth$01.05MLower standard workload estimate from another provider.Open model Compare
🔥Laguna S 2.1 (free)Poolside$0262.14KLower standard workload estimate from another provider.Open model Compare
NewDots3-Note Preview (free)Dots Studio$0512KLower standard workload estimate from another provider.Open model Compare
NewLFM2.5-2.6B (free)LiquidAI$065.54KLower standard workload estimate from another provider.Open model Compare

Same-provider swaps

Lower-cost options from the same provider, useful when account setup or procurement is already fixed.

ModelProviderSample CostContextWhy it is hereNext Step
🔥Nemotron 3 Ultra (free)NVIDIA$01MSame provider with a lower standard workload estimate.Open model Compare
New🔥Nemotron 3.5 Lightning (free)NVIDIA$01MSame provider with a lower standard workload estimate.Open model Compare
Nemotron 3.5 Content Safety (free)NVIDIA$0128KSame provider with a lower standard workload estimate.Open model Compare
Nemotron 3 Nano Omni (free)NVIDIA$0256KSame provider with a lower standard workload estimate.Open model Compare

Larger-context upgrades near this budget

Models with more context that stay within a close sample-cost band when price data is available.

ModelProviderSample CostContextWhy it is hereNext Step
🔥DeepSeek V4 Flash 0731DeepSeek$0.171.31MMore context while staying near this model's sample-cost band.Open model Compare
DeepSeek V4 Flash Latestdeepseek$0.111.31MMore context while staying near this model's sample-cost band.Open model Compare
Llama 4 ScoutMeta$0.251.31MMore context while staying near this model's sample-cost band.Open model Compare
Owl AlphaOpenRouter$01.05MMore context while staying near this model's sample-cost band.Open model Compare
Model Introduction

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Best Fit

Nemotron 3.5 Lightning is best suited for cost-sensitive production traffic.

Cost Example

A 1M input token plus 500K output token workload is estimated at $0.18.

Decision Shortcuts

Compare this model

Search head-to-head pages that include Nemotron 3.5 Lightning and review input price, output price, context, and sample workload cost.

Find comparisons

NVIDIA catalog

See other NVIDIA models before narrowing your shortlist.

Open provider hub

Cheaper alternatives

Start from models sorted by a standard cost estimate when budget is the first constraint.

Browse low-cost models

High-Interest Comparisons

Search this model
ComparisonCost-first PickContext Pick
No high-interest comparison is currently available for this model.

Popular Comparisons

Search all comparisons
ComparisonNewest Release
🔥Claude Opus 5 vs NewNemotron 3.5 Lightning
🔥Claude Sonnet 5 vs NewNemotron 3.5 Lightning
🔥DeepSeek V4 Flash 0423 vs NewNemotron 3.5 Lightning
🔥DeepSeek V4 Flash 0731 vs NewNemotron 3.5 Lightning
NewDeepSeek V4 Flash Vision Exp vs NewNemotron 3.5 Lightning
🔥DeepSeek V4 Pro 0423 vs NewNemotron 3.5 Lightning
NewDeepSeek V4 Pro 0813 vs NewNemotron 3.5 Lightning
NewDots3-Note Preview (free) vs NewNemotron 3.5 Lightning
🔥Gemini 3 Flash Preview vs NewNemotron 3.5 Lightning
🔥Gemini 3.6 Flash vs NewNemotron 3.5 Lightning