Qwen
·Released on Oct 28, 2025

Qwen3 Embedding 4B API Benchmarks, Pricing & Provider Data

Compare Qwen3 Embedding 4B with another model

Choose a model to open its comparison page.

Share on X
Embeddings

Qwen3 Embedding 4B API pricing covers 1 API provider, from $0.028/M to $0.028/M.

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capab...

Cost
$0.028/ 1M · 8:1 in:out
$0.0045 in · $0.024 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
32.8Ktokens
39.3 pages of text
8K128K1M4M
32.8K

Features

Technical Details

Input
Output
Total parameters
4B
Released
Oct 2025
Tokenizer
Other
Architecture
text->embeddings
Moderated
No
Supported parameters
frequency_penaltymax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstoptemperaturetop_ktop_p

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
DeepInfra
deepinfra
$0.020/M$0/M100%undefined tokens / —

Pricing Comparison

Qwen3 Embedding 4B API pricing starts at $0.028/M from Medu Chat.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
qwen3-embedding-4b
silliconflow
$0.028/M
$0.028/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3 Embedding 4B include?
LMSpeed shows Qwen3 Embedding 4B benchmark context, API price, output speed, first-token latency, and provider data across 1 providers when those signals are available.
What is the Qwen3 Embedding 4B API price?
Qwen3 Embedding 4B has pricing from 1 providers, ranging from $0.028/M to $0.028/M. Medu Chat has the lowest listed price.
What does the Qwen3 Embedding 4B API pricing table include?
The Qwen3 Embedding 4B API pricing table compares 1 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3 Embedding 4B API pricing?
Medu Chat currently has the lowest listed Qwen3 Embedding 4B price at $0.028/M across 1 providers.
Is Qwen3 Embedding 4B API free?
Qwen3 Embedding 4B does not currently have a free API tier on LMSpeed. All 1 providers charge per token.

Also known as

qwen3-embedding-4b

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation