Qwen3 235B A22B Thinking 2507 API Benchmarks, Pricing & Provider Data
Compare Qwen3 235B A22B Thinking 2507 with another model
Choose a model to open its comparison page.
Qwen3 235B A22B Thinking 2507 API pricing covers undefined API provider} other undefined API providers}}, from $0.031/M to $2.05/M. Qwen3 235B A22B Thinking 2507 free API options are available from undefined provider} other undefined providers}}.
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 235B
- Active parameters
- 22B
- Released
- Jul 2025
- Knowledge cutoff
- 2025-06-30
- Tokenizer
- Qwen3
- Architecture
- text->text
- Instruct type
- qwen3
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
3 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.230/M | $2.30/M | 100.0% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $0.300/M | $3/M | 96.8% | — | — | undefined tokens / undefined tokens |
Venice venice/fp8 | $0.450/M | $3.50/M | 92.2% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3 235B A22B Thinking 2507 API pricing across 22 providers. Prices range from $0.031/M to $2.05/M. Claw API offers the lowest rate at $0.031/M. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3-235b-a22b-thinking-2507 | default | $0.274/M | $2.74/M | — | — | — | |
L1 100% | qwen3-235b-a22b-thinking-2507 | default | $0.137/M | $1.37/M | — | — | — | |
L1 99% L2 91% | Qwen/Qwen3-235B-A22B-Thinking-2507 | diamond-glm | $0.080/M Cache read$0.040/M | $0.438/M | — | — | — | |
L1 99% L2 100% | qwen3-235b-a22b-thinking-2507 | alibaba | $0.569/M | $5.69/M | — | — | — | |
L1 100% | qwen3-235b-a22b-thinking-2507 | default | $0.548/M | $5.48/M | — | — | — | |
L1 100% | qwen3-235b-a22b-thinking-2507 | default | $0.274/M | $2.74/M | — | — | — | |
L1 99% | qwen3-235b-a22b-thinking-2507 | default | $1.53/M | $12.27/M | — | — | — | |
L1 99% | Qwen/Qwen3-235B-A22B-Thinking-2507 | default | $0.110/M | $0.600/M | — | — | — | |
L1 98% | qwen3-235b-a22b-thinking-2507 | 企业用户 | $2.00/M | $20.00/M | — | — | — | |
L1 0% | qwen3-235b-a22b-thinking-2507 | Qwen | $0.031/M | $0.307/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Gemini 2.5 Pro
gemini-2-5-pro
Google Gemini 2.5 Pro is Google advanced multimodal model with a 1M-token context window, strong STEM reasoning, and native support for images, audio, and video understanding.
Gemini 2.5 Flash
gemini-2-5-flash
Google Gemini 2.5 Flash is a fast multimodal model balancing speed and intelligence for chat, tool use, and large-context workloads at lower cost than Pro tiers.
GPT-5
gpt-5
OpenAI GPT-5 is OpenAI frontier general-purpose model with improved reasoning depth, coding reliability, and multimodal understanding for production assistants and agent workflows.
GLM-4.7
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
Frequently Asked Questions
- What benchmark data does Qwen3 235B A22B Thinking 2507 include?
- LMSpeed shows Qwen3 235B A22B Thinking 2507 benchmark context, API price, output speed, first-token latency, and provider data across 23 providers when those signals are available.
- What is the Qwen3 235B A22B Thinking 2507 API price?
- Qwen3 235B A22B Thinking 2507 has pricing from undefined provider} other undefined providers}}, ranging from $0.031/M to $2.05/M. Claw API has the lowest listed price.
- What does the Qwen3 235B A22B Thinking 2507 API pricing table include?
- The Qwen3 235B A22B Thinking 2507 API pricing table compares 23 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 235B A22B Thinking 2507 API pricing?
- Claw API currently has the lowest listed Qwen3 235B A22B Thinking 2507 price at $0.031/M across undefined provider} other undefined providers}}.
- Is Qwen3 235B A22B Thinking 2507 API free?
- Yes, Qwen3 235B A22B Thinking 2507 free API options are available through 1 providerundefined other undefined} on LMSpeed, including Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3 235B A22B Thinking 2507 free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Qwen3 235B A22B Thinking 2507: Zero API. Check each provider row before using it because free tier limits can change.
