Qwen3 Reranker 8B API Benchmarks, Pricing & Provider Data
Compare Qwen3 Reranker 8B with another model
Choose a model to open its comparison page.
Qwen3 Reranker 8B API pricing covers 6 API providers, from $0.0004/M to $16.00/M.
Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Fireworks fireworks | $0/M | $0/M | 100% | — | — | undefined tokens / — |
Pricing Comparison
Compare Qwen3 Reranker 8B API pricing across 6 providers. Prices range from $0.0004/M to $16.00/M. 神马中转API offers the lowest rate at $0.0004/M.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | Qwen/Qwen3-Reranker-8B | default | $16.00/M | $16.00/M | — | — | — | |
L1 100% | Qwen3-Reranker-8B | default | $16.00/M | $16.00/M | — | — | — | |
L1 99% | qwen3-reranker-8b | silliconflow | $0.056/M | $0.056/M | — | — | — | |
L1 99% | Qwen3-Reranker-8B | default | $2.00/M | $2.00/M | — | — | — | |
L1 100% | qwen3-reranker-8b | default | $0.0006/M | $0.0006/M | — | — | — |
Alternatives & Similar Models
Qwen3 Embedding
qwen3-embedding
Alibaba Qwen3 Embedding is an embedding model in the Qwen series, optimized for text embedding and similarity tasks.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Frequently Asked Questions
- What benchmark data does Qwen3 Reranker 8B include?
- LMSpeed shows Qwen3 Reranker 8B benchmark context, API price, output speed, first-token latency, and provider data across 6 providers when those signals are available.
- What is the Qwen3 Reranker 8B API price?
- Qwen3 Reranker 8B has pricing from 6 providers, ranging from $0.0004/M to $16.00/M. 神马中转API has the lowest listed price.
- What does the Qwen3 Reranker 8B API pricing table include?
- The Qwen3 Reranker 8B API pricing table compares 6 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 Reranker 8B API pricing?
- 神马中转API currently has the lowest listed Qwen3 Reranker 8B price at $0.0004/M across 6 providers.
- Is Qwen3 Reranker 8B API free?
- Qwen3 Reranker 8B does not currently have a free API tier on LMSpeed. All 6 providers charge per token.
