Qwen
·Released on Feb 25, 2026

Qwen3.5-35B-A3B API Benchmarks, Pricing & Provider Data

Compare Qwen3.5-35B-A3B with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen3.5-35B-A3B benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.027/M.

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inf...

Quality
#100of 112
44.0
LMSpeed score
Cost
#44of 186
$0.027/ 1M · 8:1 in:out
$0.0044 in · $0.023 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
7 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents42.3Coding39.9Reasoning51.1Knowledge38Math-Multilingual41.6Multimodal45.8Instruction following51.5
#1Instruction following51.5Estimated2/4 Measured dimensions
80% interval: 39.663.4
constraint following49.6
ifeval · ifeval
structured outputPrior only
novel instruction generalization53.5
ifbench · aa_if_bench
long multiturn instructionPrior only
#2Reasoning51.1RatedGlobal rank #363/4 Measured dimensions
80% interval: 42.859.4
abstract logicPrior only
scientific causal52.2
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraints56.5
mmlu_pro · mmlu_pro
evidence verification44.7
lcr · lcr / longbench · long_bench_v2
#3Multimodal45.8Estimated2/4 Measured dimensions
80% interval: 33.857.9
perception ocr49.9
v_star · v_star
document spatialPrior only
visual reasoning41.8
mathvision · math_vision
video actionPrior only
#4Agents42.3RatedGlobal rank #563/4 Measured dimensions
80% interval: 33.651.0
planning39.1
deep_planning · gert_labs
tool use52
tau · tau2_bench
environment execution35.9
browsecomp · browse_comp / osworld · os_world_verified / terminalbench · benchlm_agentic_terminal_bench2
recovery reliabilityPrior only
#5Multilingual41.6Provisional1/4 Measured dimensions
80% interval: 25.158.0
cross language understandingPrior only
multilingual generationPrior only
reasoning transfer41.6
mmlu_prox · mmlu_pro_x
low resource robustnessPrior only
#6Coding39.9Estimated2/4 Measured dimensions
80% interval: 28.751.1
code generation47
scicode · aa_sci_code
repository engineeringPrior only
debugging testing32.8
swe_rebench · swe_rebench / swe_verified · swe_verified
tooling qualityPrior only
#7Knowledge38Provisional1/4 Measured dimensions
80% interval: 21.754.3
broad knowledgePrior only
professional knowledge38
supergpqa · super_gpqa
factualityPrior only
retrieval open bookPrior only
No data:Math

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
16.4Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Total parameters
35B
Active parameters
3B
Released
Feb 2026
Tokenizer
Qwen3
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score44.0#100 / 112
Overall score
LMSpeed rank#100
Score44.0
UpdatedSep 13, 2026
Confidence4

Speed & latency

undefined metric} other undefined metrics}}

Output speed158.3 tok/s#22 / 80Time to first token1.29 s#40 / 80
Output speed
LMSpeed rank#22
Score158.3 tok/s
UpdatedSep 13, 2026
Confidence4
Time to first token
LMSpeed rank#40
Score1.29 s
UpdatedSep 13, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$0.250/M#44 / 186Output price$2.00/M#73 / 186
Input price
LMSpeed rank#44
Score$0.250/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#73
Score$2.00/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score42.3#5680% interval33.6–51.03/4 Measured dimensionsAgentic score13.7#76 / 77Terminal-Bench 2.040.5#52 / 53
Dimensions and evidence
Planning & decomposition39.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z -1.44 · q 1.00
Tool use52
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z 0.16 · q 1.00
Environment & long-horizon execution35.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -1.45 · q 1.00
osworld · os_world_verified · z -1.49 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -1.40 · q 1.00
Recovery & completion reliabilityPrior only
Agentic score
LMSpeed rank#76
Score13.7
UpdatedSep 13, 2026
Confidence4
Terminal-Bench 2.0V3 evidence
LMSpeed rank#52
Score40.5
UpdatedSep 13, 2026
Confidence4
BrowseCompV3 evidence
LMSpeed rank#28
Score61.0
UpdatedSep 13, 2026
Confidence4
OSWorld-VerifiedV3 evidence
LMSpeed rank#25
Score54.5
UpdatedSep 13, 2026
Confidence4
Τ²-bench resultsV3 evidence
LMSpeed rank#35
Score89.2
UpdatedSep 13, 2026
Confidence4
Gert LabsV3 evidence
LMSpeed rank#48
Score29.0
UpdatedSep 13, 2026
Confidence4

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score39.980% interval28.7–51.12/4 Measured dimensionsCoding score47.5#63 / 87SWE-bench Verified69.2#43 / 49
Dimensions and evidence
Code generation47
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · aa_sci_code · z -0.57 · q 1.00
Repository engineeringPrior only
Debugging & testing32.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_rebench · swe_rebench · z -1.98 · q 1.00
swe_verified · swe_verified · z -1.36 · q 1.00
Tool-assisted development & qualityPrior only
Coding score
LMSpeed rank#63
Score47.5
UpdatedSep 13, 2026
Confidence4
SWE-bench VerifiedV3 evidence
LMSpeed rank#43
Score69.2
UpdatedSep 13, 2026
Confidence4
SWE-RebenchV3 evidence
LMSpeed rank#10
Score53.7
UpdatedSep 13, 2026
Confidence4
AA-SciCodeV3 evidence
LMSpeed rank#84
Score37.7
UpdatedSep 8, 2026
Confidence4

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score51.1#3680% interval42.8–59.43/4 Measured dimensionsMMLU-Pro85.3%#27 / 129GPQA81.9%#91 / 218
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning52.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.48 · q 1.00
gpqa · gpqa · z 0.41 · q 1.00
hle · hle · z 0.27 · q 1.00
Multi-step constraints56.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 0.67 · q 1.00
Evidence integration & verification44.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z -0.09 · q 1.00
longbench · long_bench_v2 · z -1.02 · q 1.00
MMLU-ProV3 evidence
LMSpeed rank#27
Score85.3%
UpdatedSep 13, 2026
Confidence4
GPQAV3 evidence
LMSpeed rank#91
Score81.9%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#108
Score13.4%
UpdatedSep 13, 2026
Confidence4
Reasoning score
LMSpeed rank#65
Score40.2
UpdatedSep 13, 2026
Confidence4
LongBench v2V3 evidence
LMSpeed rank#9
Score59.0
UpdatedSep 13, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#59
Score72.0
UpdatedSep 13, 2026
Confidence4
CritPtV3 evidence
LMSpeed rank#74
Score0.9
UpdatedSep 13, 2026
Confidence4

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score3880% interval21.7–54.31/4 Measured dimensionsKnowledge score70.4#40 / 83SuperGPQA63.4#15 / 16
Dimensions and evidence
Broad knowledgePrior only
Professional knowledge38
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
supergpqa · super_gpqa · z -1.34 · q 1.00
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#40
Score70.4
UpdatedSep 13, 2026
Confidence4
SuperGPQAV3 evidence
LMSpeed rank#15
Score63.4
UpdatedSep 13, 2026
Confidence4
Artificial Analysis Intelligence Index
LMSpeed rank#86
Score19.3
UpdatedSep 13, 2026
Confidence4
AA-GPQA Diamond
LMSpeed rank#61
Score84.5
UpdatedSep 13, 2026
Confidence4
AA-HLE
LMSpeed rank#70
Score21.0
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Accuracy
LMSpeed rank#89
Score20.1
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Hallucination Rate
LMSpeed rank#29
Score85.4
UpdatedSep 13, 2026
Confidence4

Math

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofsPrior only
Applied & tool-assisted mathPrior only

Multilingual

V3.0

undefined metric} other undefined metrics}} · Provisional

Score41.680% interval25.1–58.01/4 Measured dimensionsMultilingual score21.1#10 / 11MMLU-ProX81.0#10 / 11
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transfer41.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_prox · mmlu_pro_x · z -0.92 · q 1.00
Low-resource robustnessPrior only
Multilingual score
LMSpeed rank#10
Score21.1
UpdatedSep 13, 2026
Confidence4
MMLU-ProXV3 evidence
LMSpeed rank#10
Score81.0
UpdatedSep 13, 2026
Confidence4

Multimodal

V3.0

undefined metric} other undefined metrics}} · Estimated

Score45.880% interval33.8–57.92/4 Measured dimensionsMMMU81.4#6 / 7MMVU72.3#6 / 6
Dimensions and evidence
Perception & OCR49.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
v_star · v_star · z -0.27 · q 1.00
Document & spatial understandingPrior only
Visual reasoning41.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mathvision · math_vision · z -1.01 · q 1.00
Video & grounded actionPrior only
MMMU
LMSpeed rank#6
Score81.4
UpdatedSep 13, 2026
Confidence4
MMVU
LMSpeed rank#6
Score72.3
UpdatedSep 13, 2026
Confidence4
MathVisionV3 evidence
LMSpeed rank#11
Score83.9
UpdatedSep 13, 2026
Confidence4
V*V3 evidence
LMSpeed rank#8
Score92.7
UpdatedSep 13, 2026
Confidence4
AA-MMMU-Pro
LMSpeed rank#44
Score72.7
UpdatedSep 13, 2026
Confidence4
Multimodal Grounded score
LMSpeed rank#29
Score67.1
UpdatedSep 13, 2026
Confidence4

Instruction following

V3.0

undefined metric} other undefined metrics}} · Estimated

Score51.580% interval39.6–63.42/4 Measured dimensionsInstruction Following score88.8#22 / 52IFEval91.9#11 / 16
Dimensions and evidence
Constraint following49.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifeval · ifeval · z -0.26 · q 1.00
Structured outputPrior only
Novel-instruction generalization53.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.17 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#22
Score88.8
UpdatedSep 13, 2026
Confidence4
IFEvalV3 evidence
LMSpeed rank#11
Score91.9
UpdatedSep 13, 2026
Confidence4
AA-IFBenchV3 evidence
LMSpeed rank#28
Score72.5
UpdatedSep 13, 2026
Confidence4

OpenRouter endpoints

7 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Venice
venice
$0.313/M$1.25/M100.0%undefined tokens / undefined tokens
Parasail
parasail/fp8
$0.150/M$1/M100.0%undefined tokens / undefined tokens
Darkbloom
darkbloom/fp4
$0.080/M$0.750/M99.9%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.240/M$1.80/M99.2%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.140/M$1/M99.1%undefined tokens / undefined tokens
Alibaba
alibaba
$0.163/M$1.30/M98.8%undefined tokens / undefined tokens
AtlasCloud
atlas-cloud/fp8
$0.225/M$1.80/M97.9%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3.5-35B-A3B API pricing across 19 providers. Prices range from $0.027/M to $77.05/M. Zhongzhuan Chat offers the lowest rate at $0.027/M.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
Qwen/Qwen3.5-35B-A3B
default
$10.27/M
$10.27/M
L1
99%
L2
84%
Qwen/Qwen3.5-35B-A3B
diamond-glm
-59%$0.102/M
Cache read$0.037/M
-64%$0.730/M
L1
99%
L2
100%
qwen3.5-35b-a3b
alibaba
$0.474/M
Cache read$0.169/M
$3.80/M
L1
100%
qwen3.5-35b-a3b
Alibaba-1
-85%$0.037/M
-85%$0.297/M
L1
100%
qwen3.5-35b-a3b
default
$10.27/M
$10.27/M
L1
99%
qwen3.5-35b-a3b
Alibaba-1
-85%$0.037/M
-85%$0.297/M
L1
100%
qwen3.5-35b-a3b
default
-89%$0.027/M
-89%$0.219/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3.5-35B-A3B include?
LMSpeed shows Qwen3.5-35B-A3B benchmark context, API price, output speed, first-token latency, and provider data across 19 providers when those signals are available.
What is the Qwen3.5-35B-A3B API price?
Qwen3.5-35B-A3B has pricing from undefined provider} other undefined providers}}, ranging from $0.027/M to $77.05/M. Zhongzhuan Chat has the lowest listed price.
What does the Qwen3.5-35B-A3B API pricing table include?
The Qwen3.5-35B-A3B API pricing table compares 19 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3.5-35B-A3B API pricing?
Zhongzhuan Chat currently has the lowest listed Qwen3.5-35B-A3B price at $0.027/M across undefined provider} other undefined providers}}.
Is Qwen3.5-35B-A3B API free?
Qwen3.5-35B-A3B does not currently have a free API tier on LMSpeed. All 19 providers charge per token.

Also known as

Qwen/Qwen3.5-35B-A3Bqwen3.5-35b-a3b

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation