Qwen
·Released on Oct 14, 2025

Qwen3 VL 8B Thinking API Benchmarks, Pricing & Provider Data

Compare Qwen3 VL 8B Thinking with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen3 VL 8B Thinking API pricing covers 23 API providers, from $0.027/M to $0.191/M.

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences...

Cost
$0.027/ 1M · 8:1 in:out
$0.0044 in · $0.023 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
32.8Ktokens
8K128K1M4M
131.1K

Features

Technical Details

Input
Output
Total parameters
8B
Released
Oct 2025
Tokenizer
Qwen3
Architecture
text+image->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Alibaba
alibaba
$0.180/M$2.10/M100%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3 VL 8B Thinking API pricing across 23 providers. Prices range from $0.027/M to $0.191/M. Zhongzhuan Chat offers the lowest rate at $0.027/M.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
qwen3-vl-8b-thinking
default
$0.064/M
$0.637/M
L1
99%
qwen3-vl-8b-thinking
default
$0.064/M
$0.637/M
L1
100%
qwen3-vl-8b-thinking
default
$0.137/M
$1.37/M
L1
100%
qwen3-vl-8b-thinking
default
$0.034/M
$0.342/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3 VL 8B Thinking include?
LMSpeed shows Qwen3 VL 8B Thinking benchmark context, API price, output speed, first-token latency, and provider data across 23 providers when those signals are available.
What is the Qwen3 VL 8B Thinking API price?
Qwen3 VL 8B Thinking has pricing from 23 providers, ranging from $0.027/M to $0.191/M. Zhongzhuan Chat has the lowest listed price.
What does the Qwen3 VL 8B Thinking API pricing table include?
The Qwen3 VL 8B Thinking API pricing table compares 23 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3 VL 8B Thinking API pricing?
Zhongzhuan Chat currently has the lowest listed Qwen3 VL 8B Thinking price at $0.027/M across 23 providers.
Is Qwen3 VL 8B Thinking API free?
Qwen3 VL 8B Thinking does not currently have a free API tier on LMSpeed. All 23 providers charge per token.

Also known as

qwen/qwen3-vl-8b-thinkingqwen3-vl-8b-thinking

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation