Nemotron 3.5 Lightning API Benchmarks, Pricing & Provider Data
Compare Nemotron 3.5 Lightning with another model
Choose a model to open its comparison page.
Nemotron 3.5 Lightning benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.00000001/request. Nemotron 3.5 Lightning free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 2 / 8
- Methodology
- V3.0
#1Reasoning48.7Provisional1/4 Measured dimensions
#2Coding42.5Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 30B
- Active parameters
- 3B
- Released
- Aug 2026
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
Detailed scores
Updated: Sep 13, 2026Speed & latency
undefined metric} other undefined metrics}}
Output speed292.1 tok/s#8 / 80Time to first token0.55 s#12 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.060/M#4 / 186Output price$0.200/M#3 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Agents
V3.0undefined metric} other undefined metrics}} · No data
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score42.580% interval26.5–58.51/4 Measured dimensionsSciCode32.1%#81 / 89
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Score48.780% interval34.8–62.61/4 Measured dimensionsGPQA74.3%#127 / 218HLE10.6%#119 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
4 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
CoreWeave coreweave/bf16 | $0.100/M | $0.250/M | 100% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/bf16 | $0.080/M | $0.200/M | 99.9% | — | — | undefined tokens / undefined tokens |
Phala phala | $0.080/M | $0.200/M | 99.9% | — | — | undefined tokens / undefined tokens |
Darkbloom darkbloom/int4 | $0.065/M | $0.180/M | 94.7% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Nemotron 3.5 Lightning API pricing across 24 providers. Prices range from $0.00000001/request to $75.00/M. Future Hub offers the lowest rate at $0.00000001/request. 5 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | nvidia/nemotron-3.5-lightning:free | default | $75.00/M | $75.00/M | — | — | — | |
DeadlySignal API Free | L1 99% L2 100% | nvidia/nemotron-3.5-lightning-30b-a3b | nvidia-free | Free | Free | — | — | — |
L1 100% | nemotron-3.5-lightning-free | default | $7.50/M | $7.50/M | — | — | — | |
L1 100% | nvidia/nemotron-3.5-lightning-30b-a3b | default | $7.50/M | $7.50/M | — | — | — | |
L1 100% | nvidia/nemotron-3.5-lightning-30b-a3b | NVIDIA英伟达 | $0.00005/request | - | — | — | — | |
L1 100% | nvidia/nemotron-3.5-lightning:free | free | $0.00005/request | - | — | — | — | |
L1 99% | nemotron-3.5-lightning | default | $0.073/request | - | — | — | — | |
L1 100% | nvidia/nemotron-3.5-lightning-30b-a3b | default | $75.00/M | $75.00/M | — | — | — | |
S3AI API Free | L1 94% L2 97% | nvidia/nemotron-3.5-lightning-30b-a3b | NIM | Free | Free | — | — | — |
L1 94% L2 97% | oc/nemotron-3.5-lightning | default | $0.0000515/request | - | — | — | — | |
Dext API Free | L1 39% | nemotron-3.5-lightning-free | 公益 | Free | Free | — | — | — |
L1 100% | nemotron-3.5-lightning | default | $0.010/request | - | — | — | — | |
L1 100% | nvidia/nemotron-3.5-lightning | openrouter | -63%$0.022/M Cache read$0.011/M | -73%$0.055/M | — | — | — | |
初叶🍂Furry API Free | L1 65% | nvidia/nemotron-3.5-lightning-30b-a3b | free | Free | Free | — | — | — |
L1 0% | nvidia/nemotron-3.5-lightning-30b-a3b | default | -54%$0.027/M | -86%$0.027/M | — | — | — | |
L1 0% | nemotron-3.5-lightning | openrouter | $0.00000001/request | - | — | — | — | |
L1 0% | nemotron-3.5-lightning | user | $0.730/M | $2.19/M | — | — | — | |
L1 0% | nemotron-3.5-lightning | 无限制 | $0.730/M | $2.19/M | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
Kimi K3
kimi-k3
Kimi K3 is an ultra-large-scale, open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at navigating...
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
Frequently Asked Questions
- What benchmark data does Nemotron 3.5 Lightning include?
- LMSpeed shows Nemotron 3.5 Lightning benchmark context, API price, output speed, first-token latency, and provider data across 29 providers when those signals are available.
- What is the Nemotron 3.5 Lightning API price?
- Nemotron 3.5 Lightning has pricing from undefined provider} other undefined providers}}, ranging from $0.00000001/request to $75.00/M. Future Hub has the lowest listed price.
- What does the Nemotron 3.5 Lightning API pricing table include?
- The Nemotron 3.5 Lightning API pricing table compares 29 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Nemotron 3.5 Lightning API pricing?
- Future Hub currently has the lowest listed Nemotron 3.5 Lightning price at $0.00000001/request across undefined provider} other undefined providers}}.
- Can I compare Nemotron 3.5 Lightning API price and speed together?
- Yes. LMSpeed shows Nemotron 3.5 Lightning API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Nemotron 3.5 Lightning API free?
- Yes, Nemotron 3.5 Lightning free API options are available through 5 providerundefined other undefined} on LMSpeed, including DeadlySignal API, DeadlySignal API, 初叶🍂Furry API, S3AI API, Dext API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Nemotron 3.5 Lightning free API access?
- LMSpeed currently lists 5 free API providerundefined other undefined} for Nemotron 3.5 Lightning: DeadlySignal API, DeadlySignal API, 初叶🍂Furry API, S3AI API, Dext API. Check each provider row before using it because free tier limits can change.
