Nemotron 3 Nano 30B A3B API Benchmarks, Pricing & Provider Data
Compare Nemotron 3 Nano 30B A3B with another model
Choose a model to open its comparison page.
Nemotron 3 Nano 30B A3B API pricing covers undefined API provider} other undefined API providers}}, from $0.020/request to $0.020/request.
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 30B
- Active parameters
- 3B
- Released
- Dec 2025
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
4 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Crusoe crusoe/fp8 | $0.050/M | $0.200/M | 100.0% | — | — | undefined tokens / undefined tokens |
Novita novita/fp4 | $0.050/M | $0.200/M | 99.9% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp4 | $0.050/M | $0.200/M | 99.9% | — | — | undefined tokens / undefined tokens |
Nebius nebius/fp8 | $0.060/M | $0.240/M | 99.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Nemotron 3 Nano 30B A3B API pricing across 2 providers. Prices range from $0.020/request to $0.020/request. 91VIP API offers the lowest rate at $0.020/request.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | laohuang/nvidia/nemotron-3-nano-30b-a3b | default | $0.020/request | - | — | — | — |
Alternatives & Similar Models
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
Qwen3.5
qwen3-5
Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.
Nemotron Nano 12B 2 VL
nemotron-nano-12b-v2-vl
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Frequently Asked Questions
- What benchmark data does Nemotron 3 Nano 30B A3B include?
- LMSpeed shows Nemotron 3 Nano 30B A3B benchmark context, API price, output speed, first-token latency, and provider data across 2 providers when those signals are available.
- What is the Nemotron 3 Nano 30B A3B API price?
- Nemotron 3 Nano 30B A3B has pricing from undefined provider} other undefined providers}}, ranging from $0.020/request to $0.020/request. 91VIP API has the lowest listed price.
- What does the Nemotron 3 Nano 30B A3B API pricing table include?
- The Nemotron 3 Nano 30B A3B API pricing table compares 2 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Nemotron 3 Nano 30B A3B API pricing?
- 91VIP API currently has the lowest listed Nemotron 3 Nano 30B A3B price at $0.020/request across undefined provider} other undefined providers}}.
- Is Nemotron 3 Nano 30B A3B API free?
- Nemotron 3 Nano 30B A3B does not currently have a free API tier on LMSpeed. All 2 providers charge per token.
