Qwen3 30B A3B Thinking 2507 API Benchmarks, Pricing & Provider Data
Compare Qwen3 30B A3B Thinking 2507 with another model
Choose a model to open its comparison page.
Qwen3 30B A3B Thinking 2507 API pricing covers undefined API provider} other undefined API providers}}, from $0.030/M to $75.00/M. Qwen3 30B A3B Thinking 2507 free API options are available from undefined provider} other undefined providers}}.
Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking m...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 30B
- Active parameters
- 3B
- Released
- Aug 2025
- Knowledge cutoff
- 2025-06-30
- Tokenizer
- Qwen3
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_p
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.200/M | $2.40/M | 100% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3 30B A3B Thinking 2507 API pricing across 25 providers. Prices range from $0.030/M to $75.00/M. Jeniya AI API offers the lowest rate at $0.030/M. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3-30b-a3b-thinking-2507 | default | $0.103/M | $1.03/M | — | — | — | |
DeadlySignal API Free | L1 99% L2 92% | Qwen/Qwen3-30B-A3B-Thinking-2507 | diamond-glm | Free | Free | — | — | — |
L1 100% | qwen3-30b-a3b-thinking-2507 | default | $0.208/M | $2.05/M | — | — | — | |
L1 100% | qwen3-30b-a3b-thinking-2507 | Alibaba-1 | $0.030/M | $0.356/M | — | — | — | |
L1 100% | qwen3-30b-a3b-thinking-2507 | default | $0.103/M | $1.03/M | — | — | — | |
L1 99% | qwen3-30b-a3b-thinking-2507 | Alibaba-1 | $0.030/M | $0.356/M | — | — | — | |
L1 99% | qwen3-30b-a3b-thinking-2507 | default | $0.192/M | $2.30/M | — | — | — | |
L1 99% | Qwen/Qwen3-30B-A3B-Thinking-2507 | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | qwen3-30b-a3b-thinking-2507 | default | $0.051/M | $0.514/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
Gemini 2.5 Pro
gemini-2-5-pro
Google Gemini 2.5 Pro is Google advanced multimodal model with a 1M-token context window, strong STEM reasoning, and native support for images, audio, and video understanding.
Gemini 2.5 Flash
gemini-2-5-flash
Google Gemini 2.5 Flash is a fast multimodal model balancing speed and intelligence for chat, tool use, and large-context workloads at lower cost than Pro tiers.
GPT-5
gpt-5
OpenAI GPT-5 is OpenAI frontier general-purpose model with improved reasoning depth, coding reliability, and multimodal understanding for production assistants and agent workflows.
Frequently Asked Questions
- What benchmark data does Qwen3 30B A3B Thinking 2507 include?
- LMSpeed shows Qwen3 30B A3B Thinking 2507 benchmark context, API price, output speed, first-token latency, and provider data across 26 providers when those signals are available.
- What is the Qwen3 30B A3B Thinking 2507 API price?
- Qwen3 30B A3B Thinking 2507 has pricing from undefined provider} other undefined providers}}, ranging from $0.030/M to $75.00/M. Jeniya AI API has the lowest listed price.
- What does the Qwen3 30B A3B Thinking 2507 API pricing table include?
- The Qwen3 30B A3B Thinking 2507 API pricing table compares 26 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 30B A3B Thinking 2507 API pricing?
- Jeniya AI API currently has the lowest listed Qwen3 30B A3B Thinking 2507 price at $0.030/M across undefined provider} other undefined providers}}.
- Is Qwen3 30B A3B Thinking 2507 API free?
- Yes, Qwen3 30B A3B Thinking 2507 free API options are available through 1 providerundefined other undefined} on LMSpeed, including DeadlySignal API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3 30B A3B Thinking 2507 free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Qwen3 30B A3B Thinking 2507: DeadlySignal API. Check each provider row before using it because free tier limits can change.
