Released on Mar 10, 2026

Qwen3.5 API Benchmarks, Pricing & Provider Data

Compare Qwen3.5 with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen3.5 benchmark, API pricing, and provider data cover 331 API providers, with prices starting at $0.0050/M. Qwen3.5 free API options are available from 9 providers. The page also shows measured API speed and first-token latency.

Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.

Quality
#59of 101
54.0
LMSpeed score
Speed
#78of 78
51char/s
10.34 s
Cost
#1of 182
$0.0050/ 1M · 8:1 in:out
$0.0008 in · $0.0042 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
8 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents44.6Coding40.3Reasoning52.3Knowledge53.1Math41.4Multilingual54.6Multimodal50.2Instruction following53.7
#1Multilingual54.6Estimated2/4 Measured dimensions
80% interval: 42.466.8
cross language understanding54.9
nova · nova63
multilingual generationPrior only
reasoning transfer54.2
mmlu_prox · mmlu_pro_x
low resource robustnessPrior only
#2Instruction following53.7Estimated2/4 Measured dimensions
80% interval: 41.865.6
constraint following51.5
ifeval · ifeval
structured outputPrior only
novel instruction generalization55.8
ifbench · aa_if_bench
long multiturn instructionPrior only
#3Knowledge53.1Estimated2/4 Measured dimensions
80% interval: 41.464.7
broad knowledge57.1
c_eval · c_eval / mmlu_redux · mmlu_redux
professional knowledge49.1
supergpqa · super_gpqa
factualityPrior only
retrieval open bookPrior only
#4Reasoning52.3RatedGlobal rank #353/4 Measured dimensions
80% interval: 44.160.4
abstract logicPrior only
scientific causal41.1
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraints59.4
mmlu_pro · mmlu_pro
evidence verification56.3
ai_needle · ai_needle / lcr · lcr / longbench · long_bench_v2
#5Multimodal50.2RatedGlobal rank #64/4 Measured dimensions
80% interval: 43.756.6
perception ocr57.3
v_star · v_star
document spatial48.6
charxiv · charxiv
visual reasoning52.5
mathvision · math_vision / mmmu_pro · mmmu_pro
video action42.3
screenspot · screen_spot_pro / video_mmmu · video_mmmu
#6Agents44.6RatedGlobal rank #504/4 Measured dimensions
80% interval: 39.349.8
planning48.8
deep_planning · gert_labs
tool use41.9
mcp_atlas · mcp_atlas / mcp_tasks · mcp_tasks / tau · tau2_bench / tau · tau3_bench / toolathlon · toolathlon
environment execution45.2
browsecomp · browse_comp / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2 / vita_bench · vita_bench / wide_research · wide_research
recovery reliability42.5
apex_agents · apex_agents_aa / claw_eval · claw_eval / researchclaw · research_claw_bench
#7Math41.4Estimated2/4 Measured dimensions
80% interval: 30.052.9
foundational mathPrior only
competition math41.2
aime · aime2026 / hmmt · hmmt_feb2025 / hmmt · hmmt_feb2026 / hmmt · hmmt_nov2025
proof frontierPrior only
applied tool math41.7
mmanswer · mm_answer_bench
#8Coding40.3RatedGlobal rank #373/4 Measured dimensions
80% interval: 31.449.3
code generation30.4
livecodebench · live_code_bench_v6 / scicode · scicode
repository engineering41.3
swe_pro · swe_pro
debugging testing49.3
swe_verified · swe_verified
tooling qualityPrior only

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
262.1Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Total parameters
9B
Released
Mar 2026
Tokenizer
Qwen3
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 2, 2026

Overall

undefined metric} other undefined metrics}}

Overall score54.0#59 / 101
Overall score
LMSpeed rank#59
Score54.0
UpdatedSep 2, 2026
Confidence4

Speed & latency

undefined metric} other undefined metrics}}

Output speed20.7 tok/s#78 / 78Time to first token0.45 s#7 / 78
Output speed
LMSpeed rank#78
Score20.7 tok/s
UpdatedSep 2, 2026
Confidence4
Time to first token
LMSpeed rank#7
Score0.45 s
UpdatedSep 2, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$0.030/M#1 / 182Output price$0.150/M#2 / 182
Input price
LMSpeed rank#1
Score$0.030/M
UpdatedSep 2, 2026
Confidence4
Output price
LMSpeed rank#2
Score$0.150/M
UpdatedSep 2, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score44.6#5080% interval39.3–49.84/4 Measured dimensionsAgentic score32.4#60 / 70Terminal-Bench 2.052.5#42 / 53
Dimensions and evidence
Planning & decomposition48.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z -0.14 · q 1.00
Tool use41.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z -2.04 · q 1.00
mcp_tasks · mcp_tasks · z 0.27 · q 0.63
tau · tau2_bench · z 0.69 · q 1.00
tau · tau3_bench · z -0.70 · q 1.00
toolathlon · toolathlon · z -1.22 · q 1.00
Environment & long-horizon execution45.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -1.27 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z -0.55 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.69 · q 1.00
vita_bench · vita_bench · z 0.37 · q 1.00
wide_research · wide_research · z 0.00 · q 1.00
Recovery & completion reliability42.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z -0.76 · q 1.00
claw_eval · claw_eval · z -0.45 · q 1.00
researchclaw · research_claw_bench · z -1.29 · q 1.00
Agentic score
LMSpeed rank#60
Score32.4
UpdatedSep 2, 2026
Confidence4
Terminal-Bench 2.0V3 evidence
LMSpeed rank#42
Score52.5
UpdatedSep 2, 2026
Confidence4
BrowseCompV3 evidence
LMSpeed rank#26
Score62.0
UpdatedSep 2, 2026
Confidence4
Claw-EvalV3 evidence
LMSpeed rank#20
Score56.8
UpdatedSep 2, 2026
Confidence4
QwenClawBench
LMSpeed rank#10
Score51.8
UpdatedSep 2, 2026
Confidence4
Τ³-bench resultsV3 evidence
LMSpeed rank#6
Score68.4
UpdatedSep 2, 2026
Confidence4
VITA-BenchV3 evidence
LMSpeed rank#4
Score43.7
UpdatedSep 2, 2026
Confidence4
DeepPlanningV3 evidence
LMSpeed rank#3
Score37.6
UpdatedSep 2, 2026
Confidence4
ToolathlonV3 evidence
LMSpeed rank#21
Score36.3
UpdatedSep 2, 2026
Confidence4
MCP AtlasV3 evidence
LMSpeed rank#28
Score46.1
UpdatedSep 2, 2026
Confidence4
MCP-TasksV3 evidence
LMSpeed rank#1
Score74.2
UpdatedSep 2, 2026
Confidence4
WideResearchV3 evidence
LMSpeed rank#5
Score74.0
UpdatedSep 2, 2026
Confidence4
Τ²-bench resultsV3 evidence
LMSpeed rank#17
Score95.6
UpdatedSep 2, 2026
Confidence4
Gert LabsV3 evidence
LMSpeed rank#27
Score46.8
UpdatedSep 2, 2026
Confidence4
ResearchClawBenchV3 evidence
LMSpeed rank#15
Score14.2
UpdatedSep 2, 2026
Confidence4
AA Agentic Index
LMSpeed rank#52
Score19.9
UpdatedSep 2, 2026
Confidence4
APEX-Agents-AAV3 evidence
LMSpeed rank#16
Score15.3
UpdatedSep 2, 2026
Confidence4
GDPval-AA
LMSpeed rank#49
Score23.3
UpdatedSep 2, 2026
Confidence4
GDPval-AAV3 evidence
LMSpeed rank#49
Score966.0
UpdatedSep 2, 2026
Confidence4

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score40.3#3780% interval31.4–49.33/4 Measured dimensionsSciCode16.1%#204 / 207Coding score46.0#63 / 82
Dimensions and evidence
Code generation30.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
livecodebench · live_code_bench_v6 · z -0.71 · q 0.75
scicode · scicode · z -3.00 · q 1.00
Repository engineering41.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_pro · swe_pro · z -1.08 · q 1.00
Debugging & testing49.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_verified · swe_verified · z -0.12 · q 1.00
Tool-assisted development & qualityPrior only
SciCodeV3 evidence
LMSpeed rank#204
Score16.1%
UpdatedSep 2, 2026
Confidence4
Coding score
LMSpeed rank#63
Score46.0
UpdatedSep 2, 2026
Confidence4
SWE-bench VerifiedV3 evidence
LMSpeed rank#29
Score76.2
UpdatedSep 2, 2026
Confidence4
LiveCodeBench v6V3 evidence
LMSpeed rank#6
Score83.6
UpdatedSep 2, 2026
Confidence4
SWE-bench ProV3 evidence
LMSpeed rank#46
Score50.9
UpdatedSep 2, 2026
Confidence4
AA-SciCodeV3 evidence
LMSpeed rank#60
Score42.0
UpdatedSep 2, 2026
Confidence4
AA Coding Index
LMSpeed rank#46
Score48.2
UpdatedSep 2, 2026
Confidence4

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score52.3#3580% interval44.1–60.43/4 Measured dimensionsMMLU-Pro87.8%#8 / 129GPQA77.1%#109 / 214
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning41.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.21 · q 1.00
gpqa · gpqa · z -1.33 · q 1.00
hle · hle · z -1.17 · q 1.00
Multi-step constraints59.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 1.05 · q 1.00
Evidence integration & verification56.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ai_needle · ai_needle · z 0.09 · q 0.50
lcr · lcr · z 0.16 · q 1.00
longbench · long_bench_v2 · z 1.14 · q 1.00
MMLU-ProV3 evidence
LMSpeed rank#8
Score87.8%
UpdatedSep 2, 2026
Confidence4
GPQAV3 evidence
LMSpeed rank#109
Score77.1%
UpdatedSep 2, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#118
Score9.9%
UpdatedSep 2, 2026
Confidence4
Reasoning score
LMSpeed rank#15
Score65.7
UpdatedSep 2, 2026
Confidence4
LongBench v2V3 evidence
LMSpeed rank#3
Score63.2
UpdatedSep 2, 2026
Confidence4
AI-NeedleV3 evidence
LMSpeed rank#2
Score68.7
UpdatedSep 2, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#44
Score72.7
UpdatedSep 2, 2026
Confidence4
CritPtV3 evidence
LMSpeed rank#59
Score1.7
UpdatedSep 2, 2026
Confidence4

Knowledge

V3.0

undefined metric} other undefined metrics}} · Estimated

Score53.180% interval41.4–64.72/4 Measured dimensionsKnowledge score53.5#57 / 70SuperGPQA70.4#8 / 16
Dimensions and evidence
Broad knowledge57.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
c_eval · c_eval · z 0.71 · q 0.63
mmlu_redux · mmlu_redux · z 0.56 · q 0.75
Professional knowledge49.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
supergpqa · super_gpqa · z 0.13 · q 1.00
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#57
Score53.5
UpdatedSep 2, 2026
Confidence4
SuperGPQAV3 evidence
LMSpeed rank#8
Score70.4
UpdatedSep 2, 2026
Confidence4
MMLU-ReduxV3 evidence
LMSpeed rank#3
Score94.9
UpdatedSep 2, 2026
Confidence4
C-EvalV3 evidence
LMSpeed rank#2
Score93.0
UpdatedSep 2, 2026
Confidence4
Artificial Analysis Intelligence Index
LMSpeed rank#66
Score34.3
UpdatedSep 2, 2026
Confidence4
AA-GPQA Diamond
LMSpeed rank#38
Score89.3
UpdatedSep 2, 2026
Confidence4
AA-HLE
LMSpeed rank#47
Score29.0
UpdatedSep 2, 2026
Confidence4
AA-Omniscience Accuracy
LMSpeed rank#51
Score30.8
UpdatedSep 2, 2026
Confidence4
AA-Omniscience Hallucination Rate
LMSpeed rank#23
Score88.9
UpdatedSep 2, 2026
Confidence4

Math

V3.0

undefined metric} other undefined metrics}} · Estimated

Score41.480% interval30.0–52.92/4 Measured dimensionsMath score74.3#14 / 62AIME2693.3#12 / 14
Dimensions and evidence
Foundational mathPrior only
Competition math41.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aime · aime2026 · z -1.54 · q 1.00
hmmt · hmmt_feb2025 · z 0.00 · q 0.88
hmmt · hmmt_feb2026 · z 0.07 · q 1.00
hmmt · hmmt_nov2025 · z -0.25 · q 1.00
Advanced proofsPrior only
Applied & tool-assisted math41.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmanswer · mm_answer_bench · z -1.06 · q 1.00
Math score
LMSpeed rank#14
Score74.3
UpdatedSep 2, 2026
Confidence4
AIME26V3 evidence
LMSpeed rank#12
Score93.3
UpdatedSep 2, 2026
Confidence4
HMMT Feb 2025V3 evidence
LMSpeed rank#4
Score94.8
UpdatedSep 2, 2026
Confidence4
HMMT Nov 2025V3 evidence
LMSpeed rank#6
Score92.7
UpdatedSep 2, 2026
Confidence4
HMMT Feb 2026V3 evidence
LMSpeed rank#10
Score87.9
UpdatedSep 2, 2026
Confidence4
MMAnswerBenchV3 evidence
LMSpeed rank#8
Score80.9
UpdatedSep 2, 2026
Confidence4

Multilingual

V3.0

undefined metric} other undefined metrics}} · Estimated

Score54.680% interval42.4–66.82/4 Measured dimensionsMultilingual score69.7#4 / 11MMLU-ProX84.7#4 / 11
Dimensions and evidence
Cross-language understanding54.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
nova · nova63 · z 0.64 · q 0.88
Multilingual generationPrior only
Reasoning transfer54.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_prox · mmlu_pro_x · z 0.77 · q 1.00
Low-resource robustnessPrior only
Multilingual score
LMSpeed rank#4
Score69.7
UpdatedSep 2, 2026
Confidence4
MMLU-ProXV3 evidence
LMSpeed rank#4
Score84.7
UpdatedSep 2, 2026
Confidence4
NOVA-63V3 evidence
LMSpeed rank#1
Score59.1
UpdatedSep 2, 2026
Confidence4

Multimodal

V3.0

undefined metric} other undefined metrics}} · Rated

Score50.2#680% interval43.7–56.64/4 Measured dimensionsMultimodal Grounded score60.9#27 / 44MMMU-Pro79.0#13 / 31
Dimensions and evidence
Perception & OCR57.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
v_star · v_star · z 0.72 · q 1.00
Document & spatial understanding48.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
charxiv · charxiv · z -0.07 · q 1.00
Visual reasoning52.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mathvision · math_vision · z 0.40 · q 1.00
mmmu_pro · mmmu_pro · z 0.12 · q 1.00
Video & grounded action42.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
screenspot · screen_spot_pro · z -1.12 · q 1.00
video_mmmu · video_mmmu · z 0.03 · q 1.00
Multimodal Grounded score
LMSpeed rank#27
Score60.9
UpdatedSep 2, 2026
Confidence4
MMMU-ProV3 evidence
LMSpeed rank#13
Score79.0
UpdatedSep 2, 2026
Confidence4
MathVisionV3 evidence
LMSpeed rank#5
Score88.6
UpdatedSep 2, 2026
Confidence4
CharXivV3 evidence
LMSpeed rank#18
Score80.8
UpdatedSep 2, 2026
Confidence4
VideoMMMUV3 evidence
LMSpeed rank#5
Score84.7
UpdatedSep 2, 2026
Confidence4
ScreenSpot ProV3 evidence
LMSpeed rank#10
Score65.6
UpdatedSep 2, 2026
Confidence4
V*V3 evidence
LMSpeed rank#3
Score95.8
UpdatedSep 2, 2026
Confidence4
AA-MMMU-Pro
LMSpeed rank#24
Score77.3
UpdatedSep 2, 2026
Confidence4

Instruction following

V3.0

undefined metric} other undefined metrics}} · Estimated

Score53.780% interval41.8–65.62/4 Measured dimensionsInstruction Following score87.4#12 / 25IFEval92.6#8 / 16
Dimensions and evidence
Constraint following51.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifeval · ifeval · z 0.00 · q 1.00
Structured outputPrior only
Novel-instruction generalization55.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.50 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#12
Score87.4
UpdatedSep 2, 2026
Confidence4
IFEvalV3 evidence
LMSpeed rank#8
Score92.6
UpdatedSep 2, 2026
Confidence4
AA-IFBenchV3 evidence
LMSpeed rank#5
Score78.8
UpdatedSep 2, 2026
Confidence4

Pricing Comparison

Compare Qwen3.5 API pricing across 322 providers. Prices range from $0.0050/M to $1718.28/M. RenRen API offers the lowest rate at $0.0050/M. 9 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
qwen3.5-35b-a3b
default
-9%$0.027/M
$0.219/M
L1
100%
qwen3.5-27b
default
$0.041/M
$0.329/M
L1
100%
qwen3.5-122b-a10b
default
$0.055/M
$0.438/M
L1
100%
qwen3.5-397b-a17b
default
$0.082/M
$0.493/M
L1
100%
Qwen3.5-35B-A3B
default
$0.020/request
-
L1
100%
Qwen3.5-122B-A10B
default
$0.020/request
-
L1
100%
qwen3.5-35b-a3b
qwen
$0.160/M
$1.28/M
L1
100%
qwen3.5-27b
qwen
$0.240/M
$1.92/M
L1
100%
qwen3.5-122b-a10b
qwen
$0.320/M
$2.56/M
L1
100%
qwen3.5-397b-a17b
qwen
$0.480/M
$2.88/M
L1
100%
qwen3.5-35b-a3b
default
$0.055/M
$0.438/M
L1
100%
qwen3.5-27b
default
$0.082/M
$0.658/M
L1
100%
qwen3.5-122b-a10b
default
$0.110/M
$0.877/M
L1
100%
qwen3.5-397b-a17b
default
$0.164/M
$0.986/M
L1
100%
qwen3.5-397b-a17b
default
$10.27/M
$10.27/M
L1
100%
qwen3.5-35b-a3b
default
-9%$0.027/M
$0.219/M
L1
100%
qwen3.5-27b
default
$0.041/M
$0.329/M
L1
100%
qwen3.5-122b-a10b
default
$0.055/M
$0.438/M
L1
100%
qwen3.5-397b-a17b
default
$0.082/M
$0.493/M
L1
100%
L2
62%
qwen3.5-122b-a10b
nvidia
$0.400/M
$3.20/M
Showing 20 model IDs of 131.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3.5 include?
LMSpeed shows Qwen3.5 benchmark context, API price, output speed, first-token latency, and provider data across 331 providers when those signals are available.
What is the Qwen3.5 API price?
Qwen3.5 has pricing from 331 providers, ranging from $0.0050/M to $1718.28/M. RenRen API has the lowest listed price.
What does the Qwen3.5 API pricing table include?
The Qwen3.5 API pricing table compares 331 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3.5 API pricing?
RenRen API currently has the lowest listed Qwen3.5 price at $0.0050/M across 331 providers.
Can I compare Qwen3.5 API price and speed together?
Yes. LMSpeed shows Qwen3.5 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Qwen3.5 API free?
Yes, Qwen3.5 free API options are available through 9 providerundefined other undefined} on LMSpeed, including Dext API, ApiToken Online, Dext API, Zero API, Zero API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Qwen3.5 free API access?
LMSpeed currently lists 9 free API providerundefined other undefined} for Qwen3.5: Dext API, ApiToken Online, Dext API, Zero API, Zero API. Check each provider row before using it because free tier limits can change.

Also known as

Qwen/Qwen3.5-122B-A10BQwen/Qwen3.5-27BQwen/Qwen3.5-35B-A3BQwen/Qwen3.5-397B-A17BQwen/Qwen3.5-4B

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation