Qwen3 Embedding 4B API Benchmarks, Pricing & Provider Data
Compare Qwen3 Embedding 4B with another model
Choose a model to open its comparison page.
Qwen3 Embedding 4B API pricing covers 1 API provider, from $0.028/M to $0.028/M.
The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capab...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
DeepInfra deepinfra | $0.020/M | $0/M | 100% | — | — | undefined tokens / — |
Pricing Comparison
Qwen3 Embedding 4B API pricing starts at $0.028/M from Medu Chat.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 99% | qwen3-embedding-4b | silliconflow | $0.028/M | $0.028/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
Frequently Asked Questions
- What benchmark data does Qwen3 Embedding 4B include?
- LMSpeed shows Qwen3 Embedding 4B benchmark context, API price, output speed, first-token latency, and provider data across 1 providers when those signals are available.
- What is the Qwen3 Embedding 4B API price?
- Qwen3 Embedding 4B has pricing from 1 providers, ranging from $0.028/M to $0.028/M. Medu Chat has the lowest listed price.
- What does the Qwen3 Embedding 4B API pricing table include?
- The Qwen3 Embedding 4B API pricing table compares 1 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 Embedding 4B API pricing?
- Medu Chat currently has the lowest listed Qwen3 Embedding 4B price at $0.028/M across 1 providers.
- Is Qwen3 Embedding 4B API free?
- Qwen3 Embedding 4B does not currently have a free API tier on LMSpeed. All 1 providers charge per token.
