NVIDIA
·Released on Dec 14, 2025

Nemotron 3 Nano 30B A3B API Benchmarks, Pricing & Provider Data

Compare Nemotron 3 Nano 30B A3B with another model

Choose a model to open its comparison page.

Share on X
LLM

Nemotron 3 Nano 30B A3B API pricing covers undefined API provider} other undefined API providers}}, from $0.020/request to $0.020/request.

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

Cost
$0.020/ 1M · 8:1 in:out
$0.0032 in · $0.017 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
235.9Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Total parameters
30B
Active parameters
3B
Released
Dec 2025
Tokenizer
Other
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

OpenRouter endpoints

4 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Crusoe
crusoe/fp8
$0.050/M$0.200/M100.0%undefined tokens / undefined tokens
Novita
novita/fp4
$0.050/M$0.200/M99.9%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp4
$0.050/M$0.200/M99.9%undefined tokens / undefined tokens
Nebius
nebius/fp8
$0.060/M$0.240/M99.0%undefined tokens / undefined tokens

Pricing Comparison

Compare Nemotron 3 Nano 30B A3B API pricing across 2 providers. Prices range from $0.020/request to $0.020/request. 91VIP API offers the lowest rate at $0.020/request.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
laohuang/nvidia/nemotron-3-nano-30b-a3b
default
$0.020/request
-

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Nemotron 3 Nano 30B A3B include?
LMSpeed shows Nemotron 3 Nano 30B A3B benchmark context, API price, output speed, first-token latency, and provider data across 2 providers when those signals are available.
What is the Nemotron 3 Nano 30B A3B API price?
Nemotron 3 Nano 30B A3B has pricing from undefined provider} other undefined providers}}, ranging from $0.020/request to $0.020/request. 91VIP API has the lowest listed price.
What does the Nemotron 3 Nano 30B A3B API pricing table include?
The Nemotron 3 Nano 30B A3B API pricing table compares 2 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Nemotron 3 Nano 30B A3B API pricing?
91VIP API currently has the lowest listed Nemotron 3 Nano 30B A3B price at $0.020/request across undefined provider} other undefined providers}}.
Is Nemotron 3 Nano 30B A3B API free?
Nemotron 3 Nano 30B A3B does not currently have a free API tier on LMSpeed. All 2 providers charge per token.

Also known as

laohuang/nvidia/nemotron-3-nano-30b-a3bnemotron-3-nano-30b-a3bnvidia/nemotron-3-nano-30b-a3bnvidia/nemotron-3-nano-30b-a3b:freeor/nvidia/nemotron-3-nano-30b-a3b:free

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation