Qwen3 VL 235B A22B Thinking API Benchmarks, Pricing & Provider Data
Compare Qwen3 VL 235B A22B Thinking with another model
Choose a model to open its comparison page.
Qwen3 VL 235B A22B Thinking API pricing covers undefined API provider} other undefined API providers}}, from $0.055/M to $2.05/M. Qwen3 VL 235B A22B Thinking free API options are available from undefined provider} other undefined providers}}.
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STE...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 235B
- Active parameters
- 22B
- Released
- Sep 2025
- Knowledge cutoff
- 2025-03-31
- Tokenizer
- Qwen3
- Architecture
- text+image->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
2 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Novita novita/bf16 | $0.980/M | $3.95/M | 99.3% | — | — | undefined tokens / undefined tokens |
Alibaba alibaba | $0.400/M | $4/M | 87.9% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3 VL 235B A22B Thinking API pricing across 27 providers. Prices range from $0.055/M to $2.05/M. Zero API offers the lowest rate at $0.055/M. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3-vl-235b-a22b-thinking | default | $0.274/M | $2.74/M | — | — | — | |
L1 100% | qwen3-vl-235b-a22b-thinking | default | $0.055/M | $0.164/M | — | — | — | |
L1 100% | qwen3-vl-235b-a22b-thinking | default | $0.548/M | $5.48/M | — | — | — | |
L1 100% L2 100% | qwen3-vl-235b-a22b-thinking | alibaba | $0.949/M Cache read$0.474/M | $11.39/M | — | — | — | |
L1 100% | qwen3-vl-235b-a22b-thinking | Alibaba-1 | $0.059/M | $0.593/M | — | — | — | |
L1 100% | qwen3-vl-235b-a22b-thinking | default | $0.274/M | $2.74/M | — | — | — | |
L1 99% | qwen3-vl-235b-a22b-thinking | Alibaba-1 | $0.059/M | $0.593/M | — | — | — | |
L1 100% | qwen3-vl-235b-a22b-thinking | default | $0.288/M | $2.88/M | — | — | — | |
L1 99% | Qwen/Qwen3-VL-235B-A22B-Thinking | default | $0.300/M | $1.20/M | — | — | — | |
L1 100% | qwen3-vl-235b-a22b-thinking | default | $0.137/M | $1.37/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
Frequently Asked Questions
- What benchmark data does Qwen3 VL 235B A22B Thinking include?
- LMSpeed shows Qwen3 VL 235B A22B Thinking benchmark context, API price, output speed, first-token latency, and provider data across 28 providers when those signals are available.
- What is the Qwen3 VL 235B A22B Thinking API price?
- Qwen3 VL 235B A22B Thinking has pricing from undefined provider} other undefined providers}}, ranging from $0.055/M to $2.05/M. Zero API has the lowest listed price.
- What does the Qwen3 VL 235B A22B Thinking API pricing table include?
- The Qwen3 VL 235B A22B Thinking API pricing table compares 28 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 VL 235B A22B Thinking API pricing?
- Zero API currently has the lowest listed Qwen3 VL 235B A22B Thinking price at $0.055/M across undefined provider} other undefined providers}}.
- Is Qwen3 VL 235B A22B Thinking API free?
- Yes, Qwen3 VL 235B A22B Thinking free API options are available through 1 providerundefined other undefined} on LMSpeed, including Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3 VL 235B A22B Thinking free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Qwen3 VL 235B A22B Thinking: Zero API. Check each provider row before using it because free tier limits can change.
