Qwen
·Released on Oct 14, 2025

Qwen3 VL 8B Thinking API Benchmarks, Pricing & Provider Data

Compare Qwen3 VL 8B Thinking with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen3 VL 8B Thinking API pricing covers undefined API provider} other undefined API providers}}, from $0.027/M to $75.00/M.

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences...

Cost
$0.027/ 1M · 8:1 in:out
$0.0043 in · $0.022 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
32.8Ktokens
8K128K1M4M
131.1K

Features

Technical Details

Input
Output
Total parameters
8B
Released
Oct 2025
Tokenizer
Qwen3
Architecture
text+image->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Alibaba
alibaba
$0.180/M$2.10/M100%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3 VL 8B Thinking API pricing across 18 providers. Prices range from $0.027/M to $75.00/M. Jeniya AI API offers the lowest rate at $0.027/M.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
Qwen/Qwen3-VL-8B-Thinking
default
$10.27/M
$10.27/M
L1
100%
qwen3-vl-8b-thinking
default
$10.27/M
$10.27/M
L1
100%
qwen3-vl-8b-thinking
default
$0.137/M
$1.37/M
L1
100%
L2
94%
Qwen/Qwen3-VL-8B-Thinking
diamond-glm
$54.75/M
$54.75/M
L1
100%
qwen3-vl-8b-thinking
Alibaba-1
$0.027/M
$0.311/M
L1
99%
qwen3-vl-8b-thinking
Alibaba-1
$0.027/M
$0.311/M
L1
99%
Qwen/Qwen3-VL-8B-Thinking
default
$75.00/M
$75.00/M
L1
100%
qwen3-vl-8b-thinking
default
$0.034/M
$0.342/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3 VL 8B Thinking include?
LMSpeed shows Qwen3 VL 8B Thinking benchmark context, API price, output speed, first-token latency, and provider data across 18 providers when those signals are available.
What is the Qwen3 VL 8B Thinking API price?
Qwen3 VL 8B Thinking has pricing from undefined provider} other undefined providers}}, ranging from $0.027/M to $75.00/M. Jeniya AI API has the lowest listed price.
What does the Qwen3 VL 8B Thinking API pricing table include?
The Qwen3 VL 8B Thinking API pricing table compares 18 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3 VL 8B Thinking API pricing?
Jeniya AI API currently has the lowest listed Qwen3 VL 8B Thinking price at $0.027/M across undefined provider} other undefined providers}}.
Is Qwen3 VL 8B Thinking API free?
Qwen3 VL 8B Thinking does not currently have a free API tier on LMSpeed. All 18 providers charge per token.

Also known as

Qwen/Qwen3-VL-8B-Thinkingqwen/qwen3-vl-8b-thinkingqwen3-vl-8b-thinking

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation