Qwen
·Released on Feb 25, 2026

Qwen3.5-27B API Benchmarks, Pricing & Provider Data

Compare Qwen3.5-27B with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen3.5-27B benchmark, API pricing, and provider data cover 30 API providers, with prices starting at $0.033/M.

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities a...

Quality
#61of 100
54.0
LMSpeed score
Cost
#54of 180
$0.033/ 1M · 8:1 in:out
$0.0053 in · $0.028 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
7 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents45.9Coding48.9Reasoning53.6Knowledge41.4Math-Multilingual45.4Multimodal49Instruction following57.1
#1Instruction following57.1Estimated2/4 Measured dimensions
80% interval: 45.269.0
constraint following59.7
ifeval · ifeval
structured outputPrior only
novel instruction generalization54.5
ifbench · aa_if_bench
long multiturn instructionPrior only
#2Reasoning53.6RatedGlobal rank #273/4 Measured dimensions
80% interval: 45.361.9
abstract logicPrior only
scientific causal53.2
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraints57.4
mmlu_pro · mmlu_pro
evidence verification50.1
lcr · lcr / longbench · long_bench_v2
#3Multimodal49Estimated2/4 Measured dimensions
80% interval: 37.061.1
perception ocr51.9
v_star · v_star
document spatialPrior only
visual reasoning46.1
mathvision · math_vision
video actionPrior only
#4Coding48.9Estimated2/4 Measured dimensions
80% interval: 37.760.1
code generation52
scicode · scicode
repository engineeringPrior only
debugging testing45.8
swe_rebench · swe_rebench / swe_verified · swe_verified
tooling qualityPrior only
#5Agents45.9RatedGlobal rank #453/4 Measured dimensions
80% interval: 37.254.5
planning45
deep_planning · gert_labs
tool use54.6
tau · tau2_bench
environment execution37.9
browsecomp · browse_comp / osworld · os_world_verified / terminalbench · benchlm_agentic_terminal_bench2
recovery reliabilityPrior only
#6Multilingual45.4Provisional1/4 Measured dimensions
80% interval: 29.061.9
cross language understandingPrior only
multilingual generationPrior only
reasoning transfer45.4
mmlu_prox · mmlu_pro_x
low resource robustnessPrior only
#7Knowledge41.4Provisional1/4 Measured dimensions
80% interval: 25.157.7
broad knowledgePrior only
professional knowledge41.4
supergpqa · super_gpqa
factualityPrior only
retrieval open bookPrior only
No data:Math

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
65.5Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Total parameters
27B
Released
Feb 2026
Tokenizer
Qwen3
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Aug 23, 2026

Overall

undefined metric} other undefined metrics}}

Overall score54.0#61 / 100
Overall score
LMSpeed rank#61
Score54.0
UpdatedAug 23, 2026
Confidence3

Pricing

undefined metric} other undefined metrics}}

Input price$0.300/M#54 / 180Output price$2.40/M#84 / 180
Input price
LMSpeed rank#54
Score$0.300/M
UpdatedAug 23, 2026
Confidence4
Output price
LMSpeed rank#84
Score$2.40/M
UpdatedAug 23, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score45.9#4580% interval37.2–54.53/4 Measured dimensionsAgentic score14.9#67 / 69Terminal-Bench 2.041.6#50 / 53
Dimensions and evidence
Planning & decomposition45
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z -0.65 · q 1.00
Tool use54.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z 0.50 · q 1.00
Environment & long-horizon execution37.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -1.33 · q 1.00
osworld · os_world_verified · z -1.37 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -1.33 · q 1.00
Recovery & completion reliabilityPrior only
Agentic score
LMSpeed rank#67
Score14.9
UpdatedAug 23, 2026
Confidence3
Terminal-Bench 2.0V3 evidence
LMSpeed rank#50
Score41.6
UpdatedAug 23, 2026
Confidence3
BrowseCompV3 evidence
LMSpeed rank#27
Score61.0
UpdatedAug 23, 2026
Confidence3
OSWorld-VerifiedV3 evidence
LMSpeed rank#24
Score56.2
UpdatedAug 23, 2026
Confidence3
Τ²-bench resultsV3 evidence
LMSpeed rank#27
Score93.9
UpdatedAug 23, 2026
Confidence3
Gert LabsV3 evidence
LMSpeed rank#38
Score39.4
UpdatedAug 23, 2026
Confidence3

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score48.980% interval37.7–60.12/4 Measured dimensionsSciCode36.7%#113 / 204Coding score58.0#34 / 81
Dimensions and evidence
Code generation52
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.13 · q 1.00
Repository engineeringPrior only
Debugging & testing45.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_rebench · swe_rebench · z 0.00 · q 1.00
swe_verified · swe_verified · z -0.82 · q 1.00
Tool-assisted development & qualityPrior only
SciCodeV3 evidence
LMSpeed rank#113
Score36.7%
UpdatedAug 23, 2026
Confidence4
Coding score
LMSpeed rank#34
Score58.0
UpdatedAug 23, 2026
Confidence3
SWE-bench VerifiedV3 evidence
LMSpeed rank#40
Score72.4
UpdatedAug 23, 2026
Confidence3
SWE-RebenchV3 evidence
LMSpeed rank#6
Score58.9
UpdatedAug 23, 2026
Confidence3
AA-SciCodeV3 evidence
LMSpeed rank#73
Score39.5
UpdatedAug 23, 2026
Confidence3

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score53.6#2780% interval45.3–61.93/4 Measured dimensionsMMLU-Pro86.1%#22 / 129GPQA84.2%#66 / 211
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning53.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.47 · q 1.00
gpqa · gpqa · z 0.55 · q 1.00
hle · hle · z 0.44 · q 1.00
Multi-step constraints57.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 0.78 · q 1.00
Evidence integration & verification50.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.15 · q 1.00
longbench · long_bench_v2 · z -0.21 · q 1.00
MMLU-ProV3 evidence
LMSpeed rank#22
Score86.1%
UpdatedAug 23, 2026
Confidence3
GPQAV3 evidence
LMSpeed rank#66
Score84.2%
UpdatedAug 23, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#98
Score13.9%
UpdatedAug 23, 2026
Confidence4
Reasoning score
LMSpeed rank#16
Score74.5
UpdatedAug 23, 2026
Confidence3
LongBench v2V3 evidence
LMSpeed rank#7
Score60.6
UpdatedAug 23, 2026
Confidence3
AA-LCRV3 evidence
LMSpeed rank#44
Score72.3
UpdatedAug 23, 2026
Confidence3
CritPtV3 evidence
LMSpeed rank#70
Score0.9
UpdatedAug 23, 2026
Confidence3

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score41.480% interval25.1–57.71/4 Measured dimensionsKnowledge score76.8#22 / 69SuperGPQA65.6#13 / 16
Dimensions and evidence
Broad knowledgePrior only
Professional knowledge41.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
supergpqa · super_gpqa · z -0.89 · q 1.00
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#22
Score76.8
UpdatedAug 23, 2026
Confidence3
SuperGPQAV3 evidence
LMSpeed rank#13
Score65.6
UpdatedAug 23, 2026
Confidence3
Artificial Analysis Intelligence Index
LMSpeed rank#63
Score34.6
UpdatedAug 23, 2026
Confidence3
AA-GPQA Diamond
LMSpeed rank#52
Score85.8
UpdatedAug 23, 2026
Confidence3
AA-HLE
LMSpeed rank#57
Score23.9
UpdatedAug 23, 2026
Confidence3
AA-Omniscience Accuracy
LMSpeed rank#83
Score20.7
UpdatedAug 23, 2026
Confidence3
AA-Omniscience Hallucination Rate
LMSpeed rank#39
Score81.5
UpdatedAug 23, 2026
Confidence3

Math

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofsPrior only
Applied & tool-assisted mathPrior only

Multilingual

V3.0

undefined metric} other undefined metrics}} · Provisional

Score45.480% interval29.0–61.91/4 Measured dimensionsMultilingual score36.8#8 / 11MMLU-ProX82.2#8 / 11
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transfer45.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_prox · mmlu_pro_x · z -0.41 · q 1.00
Low-resource robustnessPrior only
Multilingual score
LMSpeed rank#8
Score36.8
UpdatedAug 23, 2026
Confidence3
MMLU-ProXV3 evidence
LMSpeed rank#8
Score82.2
UpdatedAug 23, 2026
Confidence3

Multimodal

V3.0

undefined metric} other undefined metrics}} · Estimated

Score4980% interval37.0–61.12/4 Measured dimensionsMMMU82.3#4 / 7MMVU73.3#4 / 5
Dimensions and evidence
Perception & OCR51.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
v_star · v_star · z 0.00 · q 1.00
Document & spatial understandingPrior only
Visual reasoning46.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mathvision · math_vision · z -0.43 · q 1.00
Video & grounded actionPrior only
MMMU
LMSpeed rank#4
Score82.3
UpdatedAug 23, 2026
Confidence3
MMVU
LMSpeed rank#4
Score73.3
UpdatedAug 23, 2026
Confidence3
MathVisionV3 evidence
LMSpeed rank#10
Score86.0
UpdatedAug 23, 2026
Confidence3
V*V3 evidence
LMSpeed rank#6
Score93.7
UpdatedAug 23, 2026
Confidence3
AA-MMMU-Pro
LMSpeed rank#31
Score75.0
UpdatedAug 23, 2026
Confidence3

Instruction following

V3.0

undefined metric} other undefined metrics}} · Estimated

Score57.180% interval45.2–69.02/4 Measured dimensionsInstruction Following score93.7#4 / 25IFEval95.0#1 / 16
Dimensions and evidence
Constraint following59.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifeval · ifeval · z 1.09 · q 1.00
Structured outputPrior only
Novel-instruction generalization54.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.33 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#4
Score93.7
UpdatedAug 23, 2026
Confidence3
IFEvalV3 evidence
LMSpeed rank#1
Score95.0
UpdatedAug 23, 2026
Confidence3
AA-IFBenchV3 evidence
LMSpeed rank#18
Score75.6
UpdatedAug 23, 2026
Confidence3

OpenRouter endpoints

6 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Novita
novita/bf16
$0.300/M$2.40/M100.0%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.250/M$2/M99.6%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.260/M$2.60/M99.5%undefined tokens / undefined tokens
AtlasCloud
atlas-cloud/fp8
$0.270/M$2.16/M99.5%undefined tokens / undefined tokens
Alibaba
alibaba
$0.195/M$1.56/M99.1%undefined tokens / undefined tokens
Phala
phala
$0.300/M$2.40/M75.2%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3.5-27B API pricing across 30 providers. Prices range from $0.033/M to $0.764/M. Zhongzhuan Chat offers the lowest rate at $0.033/M.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
qwen3.5-27b
default
-75%$0.076/M
-75%$0.612/M
L1
99%
qwen3.5-27b
default
-75%$0.076/M
-75%$0.612/M
L1
100%
qwen3.5-27b
default
-86%$0.041/M
-86%$0.329/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3.5-27B include?
LMSpeed shows Qwen3.5-27B benchmark context, API price, output speed, first-token latency, and provider data across 30 providers when those signals are available.
What is the Qwen3.5-27B API price?
Qwen3.5-27B has pricing from 30 providers, ranging from $0.033/M to $0.764/M. Zhongzhuan Chat has the lowest listed price.
What does the Qwen3.5-27B API pricing table include?
The Qwen3.5-27B API pricing table compares 30 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3.5-27B API pricing?
Zhongzhuan Chat currently has the lowest listed Qwen3.5-27B price at $0.033/M across 30 providers.
Is Qwen3.5-27B API free?
Qwen3.5-27B does not currently have a free API tier on LMSpeed. All 30 providers charge per token.

Also known as

qwen3.5-27b

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation