Nemotron Nano 9B V2 API Benchmarks, Pricing & Provider Data
Compare Nemotron Nano 9B V2 with another model
Choose a model to open its comparison page.
Nemotron Nano 9B V2 API pricing covers undefined API provider} other undefined API providers}}, from $0.010/request to $0.010/request.
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and.....
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Pricing Comparison
Nemotron Nano 9B V2 API pricing starts at $0.010/request from uglycat.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 0% | nvidia/nemotron-nano-9b-v2:free | default | $0.010/request | - | — | — | — |
Alternatives & Similar Models
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
GLM-4.5 Air
glm-4-5-air
Zhipu AI GLM-4.5 Air is a lightweight GLM-4.5 variant optimized for low-latency Chinese and English dialogue, retrieval-augmented apps, and edge deployments.
LLFM 2.5 1.2B Instruct
lfm-2-5-1-2b-instruct
LFM 2.5 1.2B Instruct is a compact language model in the LFM series, optimized for low-latency responses and efficient inference.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Claude Haiku 4.5
claude-haiku-4-5
Anthropic Claude Haiku 4.5 delivers fast, low-cost responses while retaining solid instruction following for chat, classification, and lightweight coding.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Frequently Asked Questions
- What benchmark data does Nemotron Nano 9B V2 include?
- LMSpeed shows Nemotron Nano 9B V2 benchmark context, API price, output speed, first-token latency, and provider data across 1 providers when those signals are available.
- What is the Nemotron Nano 9B V2 API price?
- Nemotron Nano 9B V2 has pricing from undefined provider} other undefined providers}}, ranging from $0.010/request to $0.010/request. uglycat has the lowest listed price.
- What does the Nemotron Nano 9B V2 API pricing table include?
- The Nemotron Nano 9B V2 API pricing table compares 1 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Nemotron Nano 9B V2 API pricing?
- uglycat currently has the lowest listed Nemotron Nano 9B V2 price at $0.010/request across undefined provider} other undefined providers}}.
- Is Nemotron Nano 9B V2 API free?
- Nemotron Nano 9B V2 does not currently have a free API tier on LMSpeed. All 1 providers charge per token.
