Qwen3.5 API Benchmarks, Pricing & Provider Data
Compare Qwen3.5 with another model
Choose a model to open its comparison page.
Qwen3.5 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.00005/request. Qwen3.5 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 8 / 8
- Methodology
- V3.0
#1Multilingual54.6Estimated2/4 Measured dimensions
#2Knowledge53.1Estimated2/4 Measured dimensions
#3Reasoning51RatedGlobal rank #373/4 Measured dimensions
#4Multimodal49.7RatedGlobal rank #64/4 Measured dimensions
#5Instruction following49.2Estimated2/4 Measured dimensions
#6Coding45.3RatedGlobal rank #353/4 Measured dimensions
#7Agents43.8RatedGlobal rank #544/4 Measured dimensions
#8Math41.4Estimated2/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 9B
- Released
- Mar 2026
- Tokenizer
- Qwen3
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score50.0#83 / 112
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed25.5 tok/s#79 / 80Time to first token0.48 s#8 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.030/M#1 / 186Output price$0.150/M#2 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score43.8#5480% interval38.6–49.14/4 Measured dimensionsAgentic score30.9#67 / 77Terminal-Bench 2.052.5#42 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score45.3#3580% interval36.4–54.23/4 Measured dimensionsCoding score34.8#81 / 87SWE-bench Verified76.2#29 / 49
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Score51#3780% interval42.8–59.13/4 Measured dimensionsMMLU-Pro87.8%#8 / 129GPQA77.1%#113 / 218
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Knowledge
V3.0undefined metric} other undefined metrics}} · Estimated
Score53.180% interval41.4–64.72/4 Measured dimensionsKnowledge score54.0#68 / 83SuperGPQA70.4#8 / 16
Knowledge
V3.0undefined metric} other undefined metrics}} · Estimated
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score41.480% interval30.0–52.92/4 Measured dimensionsMath score74.2#15 / 63AIME2693.3#12 / 14
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · Estimated
Score54.680% interval42.4–66.82/4 Measured dimensionsMultilingual score69.7#4 / 11MMLU-ProX84.7#4 / 11
Multilingual
V3.0undefined metric} other undefined metrics}} · Estimated
Multimodal
V3.0undefined metric} other undefined metrics}} · Rated
Score49.7#680% interval43.3–56.24/4 Measured dimensionsMultimodal Grounded score63.0#36 / 56MMMU-Pro79.0#13 / 31
Multimodal
V3.0undefined metric} other undefined metrics}} · Rated
Instruction following
V3.0undefined metric} other undefined metrics}} · Estimated
Score49.280% interval37.3–61.12/4 Measured dimensionsIFEval92.6#8 / 16AA-IFBench51.6#55 / 84
Instruction following
V3.0undefined metric} other undefined metrics}} · Estimated
Pricing Comparison
Compare Qwen3.5 API pricing across 204 providers. Prices range from $0.00005/request to $150.00/M. CM-API 公益站 offers the lowest rate at $0.00005/request. 3 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3.5-35b-a3b | qwen | $0.160/M | $1.28/M | — | — | — | |
L1 100% | qwen3.5-27b | qwen | $0.240/M | $1.92/M | — | — | — | |
L1 100% | qwen3.5-122b-a10b | qwen | $0.320/M | $2.56/M | — | — | — | |
L1 100% | qwen3.5-397b-a17b | qwen | $0.480/M | $2.88/M | — | — | — | |
L1 100% | qwen3.5-35b-a3b | default | -9%$0.027/M | $0.219/M | — | — | — | |
L1 100% | qwen3.5-27b | default | $0.041/M | $0.329/M | — | — | — | |
L1 100% | qwen3.5-122b-a10b | default | $0.055/M | $0.438/M | — | — | — | |
L1 100% | qwen3.5-397b-a17b | default | $0.082/M | $0.493/M | — | — | — | |
L1 100% | qwen3.5-35b-a3b | default | -9%$0.027/M | $0.219/M | — | — | — | |
L1 100% | qwen3.5-27b | default | $0.041/M | $0.329/M | — | — | — | |
L1 100% | qwen3.5-122b-a10b | default | $0.055/M | $0.438/M | — | — | — | |
L1 100% | qwen3.5-397b-a17b | default | $0.082/M | $0.493/M | — | — | — | |
L1 100% | qwen3.5-397b-a17b | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.5-35b-a3b | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.5-27b | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.5-122b-a10b | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | Qwen/Qwen3.5-4B | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.5-397b-a17b | default | $0.041/M Cache read$0.0082/M | $0.247/M | — | — | — | |
L1 100% | qwen3.5-122b-a10b | default | $0.030/M | $0.213/M | — | — | — | |
L1 100% | qwen3.5-35b-a3b | Alibaba-1 | $0.037/M | $0.297/M | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
Frequently Asked Questions
- What benchmark data does Qwen3.5 include?
- LMSpeed shows Qwen3.5 benchmark context, API price, output speed, first-token latency, and provider data across 207 providers when those signals are available.
- What is the Qwen3.5 API price?
- Qwen3.5 has pricing from undefined provider} other undefined providers}}, ranging from $0.00005/request to $150.00/M. CM-API 公益站 has the lowest listed price.
- What does the Qwen3.5 API pricing table include?
- The Qwen3.5 API pricing table compares 207 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3.5 API pricing?
- CM-API 公益站 currently has the lowest listed Qwen3.5 price at $0.00005/request across undefined provider} other undefined providers}}.
- Can I compare Qwen3.5 API price and speed together?
- Yes. LMSpeed shows Qwen3.5 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Qwen3.5 API free?
- Yes, Qwen3.5 free API options are available through 3 providerundefined other undefined} on LMSpeed, including Zero API, Zero API, 初叶🍂Furry API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3.5 free API access?
- LMSpeed currently lists 3 free API providerundefined other undefined} for Qwen3.5: Zero API, Zero API, 初叶🍂Furry API. Check each provider row before using it because free tier limits can change.
