Choose a model to open its comparison page.
Nemotron Nano 12B 2 VL API pricing covers 10 API providers, from $0.0010/request to $375.00/M. Nemotron Nano 12B 2 VL free API options are available from 1 provider.
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, c...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Nemotron Nano 12B 2 VL API pricing across 9 providers. Prices range from $0.0010/request to $375.00/M. CM-API 公益站 offers the lowest rate at $0.0010/request. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | nvidia/nemotron-nano-12b-v2-vl:free | free | $0.0010/request | - | — | — | — | |
L1 100% | nvidia/nemotron-nano-12b-v2-vl | 国产模型 | $75.00/M | $75.00/M | — | — | — | |
Dext API Free | L1 100% | nemotron-nano-12b-v2-vl | 公益 | Free | Free | — | — | — |
L1 100% | nvidia/nemotron-nano-12b-v2-vl | default | $0.200/M | $0.600/M | — | — | — | |
L1 100% | nemotron-nano-12b-v2-vl | default | $60.00/M | $60.00/M | — | — | — | |
L1 99% | nvidia/nemotron-nano-12b-v2-vl | default | $375.00/M | $375.00/M | — | — | — | |
L1 96% | nemotron-nano-12b-v2-vl | 小叶币 | $0.050/request | - | — | — | — | |
L1 100% | nvidia/nemotron-nano-12b-v2-vl:free | default | $0.010/request | - | — | — | — | |
L1 0% | nvidia/nemotron-nano-12b-v2-vl | default | $0.200/M | $0.600/M | — | — | — |
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
qwen3-5
Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
claude-haiku-4-5
Anthropic Claude Haiku 4.5 delivers fast, low-cost responses while retaining solid instruction following for chat, classification, and lightweight coding.
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.