NVIDIA
·Released on Oct 10, 2025

Llama 3.3 Nemotron Super 49B V1.5 API Benchmarks, Pricing & Provider Data

Compare Llama 3.3 Nemotron Super 49B V1.5 with another model

Choose a model to open its comparison page.

Share on X
LLM

Llama 3.3 Nemotron Super 49B V1.5 API pricing covers undefined API provider} other undefined API providers}}, from $0.020/request to $0.020/request.

Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It’s post-trained for agentic workflows (RAG,...

Cost
$0.020/ 1M · 8:1 in:out
$0.0032 in · $0.017 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
16.4Ktokens
8K128K1M4M
131.1K

Features

Technical Details

Input
Output
Total parameters
49B
Released
Oct 2025
Knowledge cutoff
2024-03-31
Tokenizer
Llama3
Architecture
text->text
Moderated
No
Expiration date
2026-07-17
Supported parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_p

Pricing Comparison

Compare Llama 3.3 Nemotron Super 49B V1.5 API pricing across 3 providers. Prices range from $0.020/request to $0.020/request. 91VIP API offers the lowest rate at $0.020/request.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
nvidia/llama-3.3-nemotron-super-49b-v1.5
default
$0.020/request
-
L1
99%
laohuang/nvidia/llama-3.3-nemotron-super-49b-v1.5
default
$0.020/request
-

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Llama 3.3 Nemotron Super 49B V1.5 include?
LMSpeed shows Llama 3.3 Nemotron Super 49B V1.5 benchmark context, API price, output speed, first-token latency, and provider data across 3 providers when those signals are available.
What is the Llama 3.3 Nemotron Super 49B V1.5 API price?
Llama 3.3 Nemotron Super 49B V1.5 has pricing from undefined provider} other undefined providers}}, ranging from $0.020/request to $0.020/request. 91VIP API has the lowest listed price.
What does the Llama 3.3 Nemotron Super 49B V1.5 API pricing table include?
The Llama 3.3 Nemotron Super 49B V1.5 API pricing table compares 3 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Llama 3.3 Nemotron Super 49B V1.5 API pricing?
91VIP API currently has the lowest listed Llama 3.3 Nemotron Super 49B V1.5 price at $0.020/request across undefined provider} other undefined providers}}.
Is Llama 3.3 Nemotron Super 49B V1.5 API free?
Llama 3.3 Nemotron Super 49B V1.5 does not currently have a free API tier on LMSpeed. All 3 providers charge per token.

Also known as

laohuang/nvidia/llama-3.3-nemotron-super-49b-v1.5llama-3.3-nemotron-super-49b-v1.5nvidia/llama-3.3-nemotron-super-49b-v1.5

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation