Qwen
·Released on Oct 28, 2025

Qwen3 Embedding 8B API Benchmarks, Pricing & Provider Data

Compare Qwen3 Embedding 8B with another model

Choose a model to open its comparison page.

Share on X
Embeddings

Qwen3 Embedding 8B API pricing covers 4 API providers, from $0.056/M to $60.00/M.

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capab...

Cost
$0.056/ 1M · 8:1 in:out
$0.0090 in · $0.047 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
32.8Ktokens
39.3 pages of text
OUTPUT
28.8Ktokens
8K128K1M4M
32.8K

Features

Technical Details

Input
Output
Total parameters
8B
Released
Oct 2025
Tokenizer
Other
Architecture
text->embeddings
Moderated
No
Supported parameters
frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstoptemperaturetop_ktop_logprobstop_p

OpenRouter endpoints

3 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Nebius
nebius
$0.010/M$0/M100.0%undefined tokens / undefined tokens
DeepInfra
deepinfra
$0.010/M$0/M100.0%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.040/M$0/M99.7%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3 Embedding 8B API pricing across 4 providers. Prices range from $0.056/M to $60.00/M. Medu Chat offers the lowest rate at $0.056/M.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
qwen3-embedding-8b
default
$60.00/M
$60.00/M
L1
99%
Qwen3-Embedding-8B
default
$0.714/M
$0.714/M
L1
16%
qwen3-embedding-8b
silliconflow
$0.056/M
$0.056/M
L1
0%
Qwen/Qwen3-Embedding-8B
酒馆模型
$1.00/M
$1.00/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3 Embedding 8B include?
LMSpeed shows Qwen3 Embedding 8B benchmark context, API price, output speed, first-token latency, and provider data across 4 providers when those signals are available.
What is the Qwen3 Embedding 8B API price?
Qwen3 Embedding 8B has pricing from 4 providers, ranging from $0.056/M to $60.00/M. Medu Chat has the lowest listed price.
What does the Qwen3 Embedding 8B API pricing table include?
The Qwen3 Embedding 8B API pricing table compares 4 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3 Embedding 8B API pricing?
Medu Chat currently has the lowest listed Qwen3 Embedding 8B price at $0.056/M across 4 providers.
Is Qwen3 Embedding 8B API free?
Qwen3 Embedding 8B does not currently have a free API tier on LMSpeed. All 4 providers charge per token.

Also known as

Qwen/Qwen3-Embedding-8BQwen3-Embedding-8Bqwen3-embedding-8b

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation