NVIDIA
·Released on 16 июл. 2026 г.

Nemotron 3 Embed 1B API Benchmarks, Pricing & Provider Data

Compare Nemotron 3 Embed 1B with another model

Choose a model to open its comparison page.

Share on X
Эмбеддинги

Цены API Nemotron 3 Embed 1B у поставщиков (11): от $0.0073/request до $547.50/M.

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retri...

Стоимость
$0.0073/ 1M · 8:1 in:out
$0.0012 in · $0.0061 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
32.8Ktokens
39.3 pages of text
OUTPUT
29.5Ktokens
8K128K1M4M
32.8K

Technical Details

Вход
Выход
Всего параметров
1B
Выпущена
Jul 2026
Документация
Токенизатор
Other
Архитектура
text->embeddings
Модерация
Нет
Поддерживаемые параметры
max_tokensseedtemperaturetop_p

Сравнение цен

Compare Nemotron 3 Embed 1B API pricing across 11 providers. Prices range from $0.0073/request to $547.50/M. Futureppo offers the lowest rate at $0.0073/request.

ПровайдерРаботоспособностьВариант моделиГруппаВходные данные ($/M)Выходные данные ($/M)Скорость (т/с)Первый токенАудит
L1
100%
nvidia/nemotron-3-embed-1b
default
$75.00/M
$75.00/M
L1
99%
nemotron-3-embed-1b
default
$547.50/M
$547.50/M
L1
69%
nvidia/nemotron-3-embed-1b
default
$75.00/M
$75.00/M
L1
100%
nemotron-3-embed-1b
default
$75.00/M
$75.00/M
L1
0%
nvidia/nemotron-3-embed-1b
default
$0.027/M
$0.027/M
L1
0%
nemotron-3-embed-1b
model
$0.0073/request
-
L1
0%
nemotron-3-embed-1b
model
$0.0073/request
-

Альтернативы и похожие модели

DeepSeek V4 Flash

deepseek-v4-flash

DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.

8 общих провайдеров

DeepSeek V4 Pro

deepseek-v4-pro

DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.

8 общих провайдеров

GPT-OSS

gpt-oss

GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.

8 общих провайдеров

Qwen3.5

qwen3-5

Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.

8 общих провайдеров

Mistral Large

mistral-large

This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

8 общих провайдеров

Nemotron 3 5 Content Safety

nemotron-3-5-content-safety

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

8 общих провайдеров

Frequently Asked Questions

What benchmark data does Nemotron 3 Embed 1B include?
LMSpeed shows Nemotron 3 Embed 1B benchmark context, API price, output speed, first-token latency, and provider data across 11 providers when those signals are available.
Сколько стоит API Nemotron 3 Embed 1B?
Цены Nemotron 3 Embed 1B у поставщиков (11): от $0.0073/request до $547.50/M. Самая низкая указанная цена — у Futureppo.
What does the Nemotron 3 Embed 1B API pricing table include?
The Nemotron 3 Embed 1B API pricing table compares 11 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
У какого провайдера самая низкая цена API Nemotron 3 Embed 1B?
Сейчас самая низкая указанная цена Nemotron 3 Embed 1B — $0.0073/request у Futureppo. Сравниваются поставщики: 11.
Является ли API Nemotron 3 Embed 1B бесплатным?
У Nemotron 3 Embed 1B в настоящее время нет бесплатного тарифа API на LMSpeed. Все 11 провайдеров взимают плату за token.

Также известна как

nemotron-3-embed-1bnvidia/nemotron-3-embed-1b

Рейтинги основаны на тестах, предоставленных сообществом, и периодических зондах работоспособности. Носит рекомендательный характер, не является официальными данными.Стандартные бенчмарки могут включать BenchLM и другие открытые источники.

Работайте с данными LMSpeed

Бесплатно

Бесплатный публичный API для цен LLM, бенчмарков и данных провайдеров

  • Актуальные цены и доступность
  • Бенчмарки скорости и задержки
  • 300+ моделей, 600+ провайдеров
  • 1 000 запросов в день
Документация API