Choose a model to open its comparison page.
Llama 3.3 Nemotron Super 49B V1.5 API pricing covers 6 API providers, from $0.050/request to $375.00/M. Llama 3.3 Nemotron Super 49B V1.5 free API options are available from 2 providers.
Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It’s post-trained for agentic workflows (RAG,...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Llama 3.3 Nemotron Super 49B V1.5 API pricing across 4 providers. Prices range from $0.050/request to $375.00/M. 初叶🍂Furry API offers the lowest rate at $0.050/request. 2 providers offer free API credits or a free tier.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
L1 100% | nvidia/llama-3.3-nemotron-super-49b-v1.5 | 国产模型 | $75.00/M | $75.00/M | — | — | — | |
Dext API Бесплатно | L1 100% | llama-3.3-nemotron-super-49b-v1.5 | 公益 | Бесплатно | Бесплатно | — | — | — |
L1 99% | nvidia/llama-3.3-nemotron-super-49b-v1.5 | default | $375.00/M | $375.00/M | — | — | — | |
L1 96% | llama-3.3-nemotron-super-49b-v1.5 | 小叶币 | $0.050/request | - | — | — | — | |
91VIP API Бесплатно | L1 95% | nvidia/llama-3.3-nemotron-super-49b-v1.5 | default | Бесплатно | Бесплатно | — | — | — |
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.