Choose a model to open its comparison page.
Nemotron 3 Embed 1B API pricing covers 3 API providers, from $2.57/M to $75.00/M. Nemotron 3 Embed 1B free API options are available from 1 provider.
NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retri...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Nemotron 3 Embed 1B API pricing across 2 providers. Prices range from $2.57/M to $75.00/M. Xinjianya API offers the lowest rate at $2.57/M. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
Dext API Free | L1 100% | nemotron-3-embed-1b | 公益 | Free | Free | — | — | — |
L1 98% | nvidia/nemotron-3-embed-1b | default | $75.00/M | $75.00/M | — | — | — | |
L1 0% | nvidia/nemotron-3-embed-1b | default | $2.57/M | $2.57/M | — | — | — |
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.