Qwen
·Released on Oct 28, 2025

Qwen3 Embedding 8B API Benchmarks, Pricing & Provider Data

Compare Qwen3 Embedding 8B with another model

Choose a model to open its comparison page.

Share on X
Embeddings

Qwen3 Embedding 8B API pricing covers undefined API provider} other undefined API providers}}, from $0.0073/M to $10273.97/M. Qwen3 Embedding 8B free API options are available from undefined provider} other undefined providers}}.

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capab...

Cost
$0.0073/ 1M · 8:1 in:out
$0.0012 in · $0.0061 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
32.8Ktokens
39.3 pages of text
OUTPUT
29.5Ktokens
8K128K1M4M
32.8K

Features

Technical Details

Input
Output
Total parameters
8B
Released
Oct 2025
Tokenizer
Other
Architecture
text->embeddings
Moderated
No
Supported parameters
frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstoptemperaturetop_ktop_logprobstop_p

OpenRouter endpoints

3 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Nebius
nebius
$0.010/M$0/M100.0%undefined tokens / undefined tokens
DeepInfra
deepinfra
$0.010/M$0/M100.0%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.040/M$0/M99.5%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3 Embedding 8B API pricing across 8 providers. Prices range from $0.0073/M to $10273.97/M. DeadlySignal API offers the lowest rate at $0.0073/M. 1 provider offers free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
Qwen/Qwen3-Embedding-8B
default
$10.27/M
$10.27/M
L1
100%
Nebius/Qwen3-Embedding-8B
default
$4.00/M
$4.00/M
L1
99%
L2
92%
Qwen/Qwen3-Embedding-8B
diamond-glm
$0.0073/M
$0/M
L1
99%
Qwen/Qwen3-Embedding-8B
default
$0.010/M
$0.010/M
L1
0%
qwen3-embedding-8b
silliconflow
$0.056/M
$0.056/M
L1
46%
accounts/fireworks/models/qwen3-embedding-8b
测试专用
$10273.97/M
$10273.97/M
L1
0%
L2
92%
Qwen3-Embedding-8B
price
Free
Free
L1
0%
Qwen/Qwen3-Embedding-8B
酒馆模型
$1.00/M
$1.00/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3 Embedding 8B include?
LMSpeed shows Qwen3 Embedding 8B benchmark context, API price, output speed, first-token latency, and provider data across 9 providers when those signals are available.
What is the Qwen3 Embedding 8B API price?
Qwen3 Embedding 8B has pricing from undefined provider} other undefined providers}}, ranging from $0.0073/M to $10273.97/M. DeadlySignal API has the lowest listed price.
What does the Qwen3 Embedding 8B API pricing table include?
The Qwen3 Embedding 8B API pricing table compares 9 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3 Embedding 8B API pricing?
DeadlySignal API currently has the lowest listed Qwen3 Embedding 8B price at $0.0073/M across undefined provider} other undefined providers}}.
Is Qwen3 Embedding 8B API free?
Yes, Qwen3 Embedding 8B free API options are available through 1 providerundefined other undefined} on LMSpeed, including 10dian-API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Qwen3 Embedding 8B free API access?
LMSpeed currently lists 1 free API providerundefined other undefined} for Qwen3 Embedding 8B: 10dian-API. Check each provider row before using it because free tier limits can change.

Also known as

Nebius/Qwen3-Embedding-8BQwen/Qwen3-Embedding-8BQwen3-Embedding-8Baccounts/fireworks/models/qwen3-embedding-8bqwen3-embedding-8b

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation