Qwen3 30B A3B Instruct 2507 API Benchmarks, Pricing & Provider Data
Compare Qwen3 30B A3B Instruct 2507 with another model
Choose a model to open its comparison page.
Qwen3 30B A3B Instruct 2507 API pricing covers undefined API provider} other undefined API providers}}, from $0.030/M to $10.27/M. Qwen3 30B A3B Instruct 2507 free API options are available from undefined provider} other undefined providers}}.
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quali...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 30B
- Active parameters
- 3B
- Released
- Jul 2025
- Knowledge cutoff
- 2025-06-30
- Tokenizer
- Qwen3
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltylogit_biaslogprobsmax_tokenspresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
5 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.130/M | $0.520/M | 100.0% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake | $0.048/M | $0.193/M | 99.4% | — | — | undefined tokens / undefined tokens |
DekaLLM dekallm | $0.090/M | $0.300/M | 99.4% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.090/M | $0.300/M | 98.8% | — | — | undefined tokens / undefined tokens |
Nebius nebius/fp8 | $0.100/M | $0.300/M | 92.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3 30B A3B Instruct 2507 API pricing across 28 providers. Prices range from $0.030/M to $10.27/M. Jeniya AI API offers the lowest rate at $0.030/M. 2 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3-30b-a3b-instruct-2507 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen/qwen3-30b-a3b-instruct-2507 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3-30b-a3b-instruct-2507 | default | $0.208/M | $0.822/M | — | — | — | |
L1 99% L2 100% | qwen3-30b-a3b-instruct-2507 | alibaba | $0.188/M | $0.568/M | — | — | — | |
L1 100% | qwen3-30b-a3b-instruct-2507 | Alibaba-1 | $0.030/M | $0.119/M | — | — | — | |
L1 100% | qwen3-30b-a3b-instruct-2507 | default | $0.103/M Cache read$0.021/M | $0.411/M | — | — | — | |
L1 99% | qwen3-30b-a3b-instruct-2507 | Alibaba-1 | $0.030/M | $0.119/M | — | — | — | |
L1 99% | qwen3-30b-a3b-instruct-2507 | default | $0.192/M | $0.767/M | — | — | — | |
L1 99% | Qwen/Qwen3-30B-A3B-Instruct-2507 | default | $0.080/M | $0.330/M | — | — | — | |
Dext API Free | L1 45% | Qwen3-30B-A3B-Instruct-2507 | 公益 | Free | Free | — | — | — |
L1 100% | qwen3-30b-a3b-instruct-2507 | default | $0.051/M | $0.205/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
Gemini 2.5 Pro
gemini-2-5-pro
Google Gemini 2.5 Pro is Google advanced multimodal model with a 1M-token context window, strong STEM reasoning, and native support for images, audio, and video understanding.
Gemini 2.5 Flash
gemini-2-5-flash
Google Gemini 2.5 Flash is a fast multimodal model balancing speed and intelligence for chat, tool use, and large-context workloads at lower cost than Pro tiers.
GPT-5
gpt-5
OpenAI GPT-5 is OpenAI frontier general-purpose model with improved reasoning depth, coding reliability, and multimodal understanding for production assistants and agent workflows.
Frequently Asked Questions
- What benchmark data does Qwen3 30B A3B Instruct 2507 include?
- LMSpeed shows Qwen3 30B A3B Instruct 2507 benchmark context, API price, output speed, first-token latency, and provider data across 30 providers when those signals are available.
- What is the Qwen3 30B A3B Instruct 2507 API price?
- Qwen3 30B A3B Instruct 2507 has pricing from undefined provider} other undefined providers}}, ranging from $0.030/M to $10.27/M. Jeniya AI API has the lowest listed price.
- What does the Qwen3 30B A3B Instruct 2507 API pricing table include?
- The Qwen3 30B A3B Instruct 2507 API pricing table compares 30 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3 30B A3B Instruct 2507 API pricing?
- Jeniya AI API currently has the lowest listed Qwen3 30B A3B Instruct 2507 price at $0.030/M across undefined provider} other undefined providers}}.
- Is Qwen3 30B A3B Instruct 2507 API free?
- Yes, Qwen3 30B A3B Instruct 2507 free API options are available through 2 providerundefined other undefined} on LMSpeed, including Dext API, Dext API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3 30B A3B Instruct 2507 free API access?
- LMSpeed currently lists 2 free API providerundefined other undefined} for Qwen3 30B A3B Instruct 2507: Dext API, Dext API. Check each provider row before using it because free tier limits can change.
