Nemotron 3 Embed 1B API Benchmarks, Pricing & Provider Data
Compare Nemotron 3 Embed 1B with another model
Choose a model to open its comparison page.
Nemotron 3 Embed 1B API pricing covers undefined API provider} other undefined API providers}}, from $0.0073/request to $547.50/M.
NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retri...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Pricing Comparison
Compare Nemotron 3 Embed 1B API pricing across 11 providers. Prices range from $0.0073/request to $547.50/M. Futureppo offers the lowest rate at $0.0073/request.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | nvidia/nemotron-3-embed-1b | default | $75.00/M | $75.00/M | — | — | — | |
L1 99% | nemotron-3-embed-1b | default | $547.50/M | $547.50/M | — | — | — | |
L1 69% | nvidia/nemotron-3-embed-1b | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | nemotron-3-embed-1b | default | $75.00/M | $75.00/M | — | — | — | |
L1 0% | nvidia/nemotron-3-embed-1b | default | $0.027/M | $0.027/M | — | — | — | |
L1 0% | nemotron-3-embed-1b | model | $0.0073/request | - | — | — | — | |
L1 0% | nemotron-3-embed-1b | model | $0.0073/request | - | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
Qwen3.5
qwen3-5
Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.
Mistral Large
mistral-large
This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
Nemotron 3 5 Content Safety
nemotron-3-5-content-safety
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Frequently Asked Questions
- What benchmark data does Nemotron 3 Embed 1B include?
- LMSpeed shows Nemotron 3 Embed 1B benchmark context, API price, output speed, first-token latency, and provider data across 11 providers when those signals are available.
- What is the Nemotron 3 Embed 1B API price?
- Nemotron 3 Embed 1B has pricing from undefined provider} other undefined providers}}, ranging from $0.0073/request to $547.50/M. Futureppo has the lowest listed price.
- What does the Nemotron 3 Embed 1B API pricing table include?
- The Nemotron 3 Embed 1B API pricing table compares 11 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Nemotron 3 Embed 1B API pricing?
- Futureppo currently has the lowest listed Nemotron 3 Embed 1B price at $0.0073/request across undefined provider} other undefined providers}}.
- Is Nemotron 3 Embed 1B API free?
- Nemotron 3 Embed 1B does not currently have a free API tier on LMSpeed. All 11 providers charge per token.
