NVIDIA: Nemotron 3.5 Lightning
nvidia/nemotron-3.5-lightning
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Source: OpenRouter (live)
- Input
- $0.070
- USD per 1M tokens
- Output
- $0.200
- USD per 1M tokens
- Blended
- $0.103
- 75% in / 25% out
- Context
- 262K
- tokens
- Max output
- 236K
- tokens
- Modality
- text->text
Prices as listed on OpenRouter, refreshed every five minutes. Provider-direct prices may differ. Convert to your currency on the AI Tracker.
Capabilities
Declared on the listing
- Tool / function calling Declared
- Reasoning controls Declared
- Structured output Declared
- InputText
- OutputText
Capabilities are read from the provider listing. “Not declared” means the listing does not advertise the feature, not that it is absent.
Why it costs this
Pricing context
- 87% below median
- Cheap input
- Cheap output
- Market median blended price is $0.806/M across 417 priced models.
Alternatives
More from NVIDIA
$0.200
/M blended
NVIDIA
$1.05
/M blended
NVIDIA
$0.172
/M blended
$0.105
/M blended
Alternatives
Similar price, other providers
Amazon
$0.105
/M blended
Rekaai
$0.100
/M blended
Mistral
$0.100
/M blended
Ibm Granite
$0.107
/M blended
Inference.net: Schematron V2 Small
Inference Net
$0.095
/M blended
Qwen (Alibaba)
$0.112
/M blended
Within ±50% of $0.103/M blended.
Nemotron open models. All NVIDIA models · Official site
Newsletter
Stay Ahead of AI
Get the most important developments in artificial intelligence delivered directly to your inbox.
No hype. No spam.
Just trusted insights, major model releases, pricing updates, benchmark changes, product reviews, and practical guidance from across the AI ecosystem.
- Weekly AI Briefing
- Major Model Releases
- Pricing & Benchmark Updates
- Unsubscribe Anytime
The AI HQ Briefing
One email a week. Read in five minutes.