Qwen3 VL 30B A3B Thinking API Benchmarks, Pricing & Provider Data
Compare Qwen3 VL 30B A3B Thinking with another model
Choose a model to open its comparison page.
Qwen3 VL 30B A3B Thinking API pricing covers undefined API provider} other undefined API providers}}, from $0.030/M to $75.00/M.
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex ...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 30B
- Active parameters
- 3B
- Released
- Oct 2025
- Knowledge cutoff
- 2025-03-31
- Tokenizer
- Qwen3
- Architecture
- text+image->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
2 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.200/M | $2.40/M | 100% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.290/M | $1/M | 99.5% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3 VL 30B A3B Thinking API pricing across 18 providers. Prices range from $0.030/M to $75.00/M. 简易-API中转站 offers the lowest rate at $0.030/M.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | Qwen/Qwen3-VL-30B-A3B-Thinking | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3-vl-30b-a3b-thinking | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3-vl-30b-a3b-thinking | default | $0.208/M | $2.05/M | — | — | — | |
L1 100% | qwen3-vl-30b-a3b-thinking | Alibaba-1 | $0.030/M | $0.356/M | — | — | — | |
L1 99% | qwen3-vl-30b-a3b-thinking | Alibaba-1 | $0.030/M | $0.356/M | — | — | — | |
L1 99% | qwen3-vl-30b-a3b-thinking | default | $0.115/M | $1.15/M | — | — | — | |
L1 99% | Qwen/Qwen3-VL-30B-A3B-Thinking | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | qwen3-vl-30b-a3b-thinking | default | $0.051/M | $0.514/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
DeepSeek V3.2
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
Frequently Asked Questions
- What benchmark data does Qwen3 VL 30B A3B Thinking include?
- LMSpeed shows Qwen3 VL 30B A3B Thinking benchmark context, API price, output speed, first-token latency, and provider data across 18 providers when those signals are available.
- What is the Qwen3 VL 30B A3B Thinking API price?
- Qwen3 VL 30B A3B Thinking has pricing from undefined provider} other undefined providers}}, ranging from $0.030/M to $75.00/M. 简易-API中转站 has the lowest listed price.
- What does the Qwen3 VL 30B A3B Thinking API pricing table include?
- The Qwen3 VL 30B A3B Thinking API pricing table compares 18 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 VL 30B A3B Thinking API pricing?
- 简易-API中转站 currently has the lowest listed Qwen3 VL 30B A3B Thinking price at $0.030/M across undefined provider} other undefined providers}}.
- Is Qwen3 VL 30B A3B Thinking API free?
- Qwen3 VL 30B A3B Thinking does not currently have a free API tier on LMSpeed. All 18 providers charge per token.
