Choose a model to open its comparison page.
Nemotron 3 Nano 30B A3B API pricing covers 8 API providers, from $0.0010/request to $375.00/M. Nemotron 3 Nano 30B A3B free API options are available from 1 provider.
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Nemotron 3 Nano 30B A3B API pricing across 7 providers. Prices range from $0.0010/request to $375.00/M. CM-API 公益站 offers the lowest rate at $0.0010/request. 1 provider offers free API credits or a free tier.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
L1 100% | nvidia/nemotron-3-nano-30b-a3b:free | free | $0.0010/request | - | — | — | — | |
L1 100% | nvidia/nemotron-3-nano-30b-a3b | 国产模型 | $75.00/M | $75.00/M | — | — | — | |
Dext API Бесплатно | L1 100% | nemotron-3-nano-30b-a3b | 公益 | Бесплатно | Бесплатно | — | — | — |
L1 100% | nvidia/nemotron-3-nano-30b-a3b | default | $0.050/M | $0.200/M | — | — | — | |
L1 99% | nvidia/nemotron-3-nano-30b-a3b | default | $375.00/M | $375.00/M | — | — | — | |
L1 96% | nemotron-3-nano-30b-a3b | 小叶币 | $0.050/request | - | — | — | — | |
L1 0% | nvidia/nemotron-3-nano-30b-a3b | default | $0.050/M | $0.200/M | — | — | — |
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
qwen3-5
Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.
nemotron-nano-12b-v2-vl
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.