NVIDIA
·Released on Feb 25, 2026

Llama Nemotron Embed VL 1B V2 API Benchmarks, Pricing & Provider Data

Compare Llama Nemotron Embed VL 1B V2 with another model

Choose a model to open its comparison page.

Share on X
Embeddings

Llama Nemotron Embed VL 1B V2 API pricing covers undefined API provider} other undefined API providers}}, from $0.0000001/request to $7.50/M. Llama Nemotron Embed VL 1B V2 free API options are available from undefined provider} other undefined providers}}.

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...

Cost
$0.0000001/ 1M · 8:1 in:out
$0.000000016 in · $0.000000084 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
118.0Ktokens
8K128K1M4M
131.1K

Technical Details

Input
Output
Total parameters
1B
Released
Feb 2026
Tokenizer
Other
Architecture
text+image->embeddings
Moderated
No
Supported parameters
max_tokensseedtemperaturetop_p

Pricing Comparison

Compare Llama Nemotron Embed VL 1B V2 API pricing across 5 providers. Prices range from $0.0000001/request to $7.50/M. 91VIP API offers the lowest rate at $0.0000001/request. 3 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
nvidia/llama-nemotron-embed-vl-1b-v2
default
$2.00/M
$2.00/M
L1
100%
nvidia/llama-nemotron-embed-vl-1b-v2
default
$7.50/M
$7.50/M
L1
100%
nvidia/llama-nemotron-embed-vl-1b-v2:free
default
$0.0000001/request
-
L1
100%
nvidia/llama-nemotron-embed-vl-1b-v2
default
$0.020/request
-
L1
76%
llama-nemotron-embed-vl-1b-v2
default
Free
Free
L1
76%
nvidia/llama-nemotron-embed-vl-1b-v2
default
Free
Free
L1
76%
nvidia/llama-nemotron-embed-vl-1b-v2:free
default
Free
Free

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Llama Nemotron Embed VL 1B V2 include?
LMSpeed shows Llama Nemotron Embed VL 1B V2 benchmark context, API price, output speed, first-token latency, and provider data across 8 providers when those signals are available.
What is the Llama Nemotron Embed VL 1B V2 API price?
Llama Nemotron Embed VL 1B V2 has pricing from undefined provider} other undefined providers}}, ranging from $0.0000001/request to $7.50/M. 91VIP API has the lowest listed price.
What does the Llama Nemotron Embed VL 1B V2 API pricing table include?
The Llama Nemotron Embed VL 1B V2 API pricing table compares 8 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Llama Nemotron Embed VL 1B V2 API pricing?
91VIP API currently has the lowest listed Llama Nemotron Embed VL 1B V2 price at $0.0000001/request across undefined provider} other undefined providers}}.
Is Llama Nemotron Embed VL 1B V2 API free?
Yes, Llama Nemotron Embed VL 1B V2 free API options are available through 3 providerundefined other undefined} on LMSpeed, including 梦德 API, 梦德 API, 梦德 API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Llama Nemotron Embed VL 1B V2 free API access?
LMSpeed currently lists 3 free API providerundefined other undefined} for Llama Nemotron Embed VL 1B V2: 梦德 API, 梦德 API, 梦德 API. Check each provider row before using it because free tier limits can change.

Also known as

llama-nemotron-embed-vl-1b-v2nvidia/llama-nemotron-embed-vl-1b-v2nvidia/llama-nemotron-embed-vl-1b-v2:free

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation