Llama Nemotron Rerank VL 1B V2 API Benchmarks, Pricing & Provider Data
Compare Llama Nemotron Rerank VL 1B V2 with another model
Choose a model to open its comparison page.
Llama Nemotron Rerank VL 1B V2 API pricing covers undefined API provider} other undefined API providers}}, from $0.0000001/request to $0.0000001/request.
Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Pricing Comparison
Compare Llama Nemotron Rerank VL 1B V2 API pricing across 2 providers. Prices range from $0.0000001/request to $0.0000001/request. 91VIP API offers the lowest rate at $0.0000001/request.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | nvidia/llama-nemotron-rerank-vl-1b-v2:free | default | $0.0000001/request | - | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Frequently Asked Questions
- What benchmark data does Llama Nemotron Rerank VL 1B V2 include?
- LMSpeed shows Llama Nemotron Rerank VL 1B V2 benchmark context, API price, output speed, first-token latency, and provider data across 2 providers when those signals are available.
- What is the Llama Nemotron Rerank VL 1B V2 API price?
- Llama Nemotron Rerank VL 1B V2 has pricing from undefined provider} other undefined providers}}, ranging from $0.0000001/request to $0.0000001/request. 91VIP API has the lowest listed price.
- What does the Llama Nemotron Rerank VL 1B V2 API pricing table include?
- The Llama Nemotron Rerank VL 1B V2 API pricing table compares 2 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Llama Nemotron Rerank VL 1B V2 API pricing?
- 91VIP API currently has the lowest listed Llama Nemotron Rerank VL 1B V2 price at $0.0000001/request across undefined provider} other undefined providers}}.
- Is Llama Nemotron Rerank VL 1B V2 API free?
- Llama Nemotron Rerank VL 1B V2 does not currently have a free API tier on LMSpeed. All 2 providers charge per token.
