Qwen3 Embedding 8B API Benchmarks, Pricing & Provider Data
Compare Qwen3 Embedding 8B with another model
Choose a model to open its comparison page.
Qwen3 Embedding 8B API pricing covers undefined API provider} other undefined API providers}}, from $0.0073/M to $10273.97/M. Qwen3 Embedding 8B free API options are available from undefined provider} other undefined providers}}.
The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capab...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
OpenRouter endpoints
3 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Nebius nebius | $0.010/M | $0/M | 100.0% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra | $0.010/M | $0/M | 100.0% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.040/M | $0/M | 99.5% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3 Embedding 8B API pricing across 8 providers. Prices range from $0.0073/M to $10273.97/M. DeadlySignal API offers the lowest rate at $0.0073/M. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | Qwen/Qwen3-Embedding-8B | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | Nebius/Qwen3-Embedding-8B | default | $4.00/M | $4.00/M | — | — | — | |
L1 99% L2 92% | Qwen/Qwen3-Embedding-8B | diamond-glm | $0.0073/M | $0/M | — | — | — | |
L1 99% | Qwen/Qwen3-Embedding-8B | default | $0.010/M | $0.010/M | — | — | — | |
L1 0% | qwen3-embedding-8b | silliconflow | $0.056/M | $0.056/M | — | — | — | |
L1 46% | accounts/fireworks/models/qwen3-embedding-8b | 测试专用 | $10273.97/M | $10273.97/M | — | — | — | |
10dian-API Free | L1 0% L2 92% | Qwen3-Embedding-8B | price | Free | Free | — | — | — |
L1 0% | Qwen/Qwen3-Embedding-8B | 酒馆模型 | $1.00/M | $1.00/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Gemini 3.1 Pro
gemini-3-1-pro
Google Gemini 3.1 Pro is a Gemini 3 series model with advanced multimodal reasoning, long-context support, and strong performance on coding and analytical tasks.
Gemini 3 Flash
gemini-3-flash
Google Gemini 3 Flash is a next-generation fast multimodal model for responsive assistants, document understanding, and high-throughput API traffic.
GPT-5.6 Terra
gpt-5-6-terra
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
GPT-5.6 Luna
gpt-5-6-luna
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Gemini 3.1 Flash Image
gemini-3-1-flash-image
Google Gemini 3.1 Flash Image is an image generation model, capable of producing images from text prompts.
Frequently Asked Questions
- What benchmark data does Qwen3 Embedding 8B include?
- LMSpeed shows Qwen3 Embedding 8B benchmark context, API price, output speed, first-token latency, and provider data across 9 providers when those signals are available.
- What is the Qwen3 Embedding 8B API price?
- Qwen3 Embedding 8B has pricing from undefined provider} other undefined providers}}, ranging from $0.0073/M to $10273.97/M. DeadlySignal API has the lowest listed price.
- What does the Qwen3 Embedding 8B API pricing table include?
- The Qwen3 Embedding 8B API pricing table compares 9 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 Embedding 8B API pricing?
- DeadlySignal API currently has the lowest listed Qwen3 Embedding 8B price at $0.0073/M across undefined provider} other undefined providers}}.
- Is Qwen3 Embedding 8B API free?
- Yes, Qwen3 Embedding 8B free API options are available through 1 providerundefined other undefined} on LMSpeed, including 10dian-API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3 Embedding 8B free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Qwen3 Embedding 8B: 10dian-API. Check each provider row before using it because free tier limits can change.
