Qwen
·Released on Aug 3, 2026

Qwen3.8 Max API Benchmarks, Pricing & Provider Data

Compare Qwen3.8 Max with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen3.8 Max benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0001/request. Qwen3.8 Max free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual ...

Quality
#6of 112
77.0
LMSpeed score
Speed
#75of 80
292char/s
4.31 s
Cost
#143of 186
$0.0001/ 1M · 8:1 in:out
$0.000016 in · $0.000084 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents61.5Coding61.7Reasoning63.5Knowledge52.2Math-Multilingual-Multimodal64.3Instruction following57.9
#1Multimodal64.3RatedGlobal rank #14/4 Measured dimensions
80% interval: 58.070.6
perception ocr54.7
simplevqa · simple_vqa
document spatial65.2
charxiv · charxiv
visual reasoning72.6
erqa · erqa / mathvision · math_vision / medxpert · med_xpert_qa_mm / mmmu_pro · mmmu_pro
video action64.6
screenspot · screen_spot_pro / video_mmmu · video_mmmu
#2Reasoning63.5Estimated2/4 Measured dimensions
80% interval: 53.373.6
abstract logicPrior only
scientific causal62.3
critpt · critpt / gpqa · gpqa / hle · hle / hle · hle_no_tools
multistep constraintsPrior only
evidence verification64.6
lcr · lcr / longbench · long_bench_v2
#3Coding61.7Estimated2/4 Measured dimensions
80% interval: 50.572.9
code generation58.6
scicode · scicode
repository engineering64.7
nl2repo · nl2_repo / swe_pro · swe_pro
debugging testingPrior only
tooling qualityPrior only
#4Agents61.5Estimated2/4 Measured dimensions
80% interval: 50.672.4
planningPrior only
tool usePrior only
environment execution64.8
gdpval_aa · benchlm_agentic_gdpval_aa / osworld · os_world2 / osworld · os_world_verified / wide_research · wide_research
recovery reliability58.2
jobbench · job_bench
#5Instruction following57.9Provisional1/4 Measured dimensions
80% interval: 41.973.9
constraint followingPrior only
structured outputPrior only
novel instruction generalization57.9
ifbench · if_bench
long multiturn instructionPrior only
#6Knowledge52.2Provisional1/4 Measured dimensions
80% interval: 38.266.2
broad knowledge52.2
aa_omniscience · aa_omniscience_index / benchlm_category_knowledge · benchlm_category_knowledge
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
No data:MathMultilingual

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1Mtokens
1.2K pages of text
OUTPUT
131.1Ktokens
8K128K1M4M
1M

Features

Technical Details

Input
Output
Released
Aug 2026
Documentation
Tokenizer
Qwen
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 14, 2026

Overall

undefined metric} other undefined metrics}}

Overall score77.0#6 / 112SkillsBench70.2#1 / 3
Overall score
LMSpeed rank#6
Score77.0
UpdatedSep 14, 2026
Confidence3
SkillsBench
LMSpeed rank#1
Score70.2
UpdatedSep 14, 2026
Confidence3
DeepSWE
LMSpeed rank#15
Score56.6
UpdatedSep 10, 2026
Confidence3

Speed & latency

undefined metric} other undefined metrics}}

Output speed43.0 tok/s#75 / 80Time to first token1.69 s#46 / 80
Output speed
LMSpeed rank#75
Score43.0 tok/s
UpdatedSep 14, 2026
Confidence4
Time to first token
LMSpeed rank#46
Score1.69 s
UpdatedSep 14, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$2.00/M#143 / 186Output price$6.00/M#128 / 186
Input price
LMSpeed rank#143
Score$2.00/M
UpdatedSep 14, 2026
Confidence4
Output price
LMSpeed rank#128
Score$6.00/M
UpdatedSep 14, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Estimated

Score61.580% interval50.6–72.42/4 Measured dimensionsAgentic score83.6#9 / 77Terminal-Bench 2.186.6#4 / 10
Dimensions and evidence
Planning & decompositionPrior only
Tool usePrior only
Environment & long-horizon execution64.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
gdpval_aa · benchlm_agentic_gdpval_aa · z 1.44 · q 1.00
osworld · os_world2 · z -0.05 · q 1.00
osworld · os_world_verified · z 1.50 · q 1.00
wide_research · wide_research · z 1.62 · q 1.00
Recovery & completion reliability58.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
jobbench · job_bench · z 0.94 · q 1.00
Agentic score
LMSpeed rank#9
Score83.6
UpdatedSep 14, 2026
Confidence3
Terminal-Bench 2.1
LMSpeed rank#4
Score86.6
UpdatedSep 14, 2026
Confidence3
CoWorkBench
LMSpeed rank#1
Score74.8
UpdatedSep 14, 2026
Confidence3
JobBenchV3 evidence
LMSpeed rank#3
Score53.4
UpdatedSep 14, 2026
Confidence3
Agents' Last Exam
LMSpeed rank#2
Score52.4
UpdatedSep 14, 2026
Confidence3
AutomationBench
LMSpeed rank#8
Score27.3
UpdatedSep 14, 2026
Confidence3
Toolathlon-Verified
LMSpeed rank#6
Score72.5
UpdatedSep 14, 2026
Confidence3
WideResearchV3 evidence
LMSpeed rank#1
Score81.9
UpdatedSep 14, 2026
Confidence3
HLE w/ tools
LMSpeed rank#5
Score56.2
UpdatedSep 14, 2026
Confidence3
OSWorld-VerifiedV3 evidence
LMSpeed rank#1
Score86.1
UpdatedSep 14, 2026
Confidence3
OSWorld 2.0V3 evidence
LMSpeed rank#11
Score19.4
UpdatedSep 14, 2026
Confidence3
WebArena-Verified
LMSpeed rank#2
Score66.8
UpdatedSep 14, 2026
Confidence3
AndroidWorld
LMSpeed rank#1
Score85.3
UpdatedSep 14, 2026
Confidence3
MobileWorld
LMSpeed rank#1
Score77.8
UpdatedSep 14, 2026
Confidence3
Terminal-Bench 2.1 (Vals)
LMSpeed rank#22
Score67.4
UpdatedSep 14, 2026
Confidence3
AA Agentic Index
LMSpeed rank#1
Score58.4
UpdatedAug 14, 2026
Confidence1
GDPval-AA
LMSpeed rank#2
Score62.0
UpdatedAug 14, 2026
Confidence1
GDPval-AAV3 evidence
LMSpeed rank#6
Score1739.0
UpdatedAug 14, 2026
Confidence1

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score61.780% interval50.5–72.92/4 Measured dimensionsSciCode53.2%#23 / 89Coding score68.0#14 / 87
Dimensions and evidence
Code generation58.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.97 · q 1.00
Repository engineering64.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
nl2repo · nl2_repo · z 1.36 · q 1.00
swe_pro · swe_pro · z 1.95 · q 1.00
Debugging & testingPrior only
Tool-assisted development & qualityPrior only
SciCodeV3 evidence
LMSpeed rank#23
Score53.2%
UpdatedSep 14, 2026
Confidence4
Coding score
LMSpeed rank#14
Score68.0
UpdatedSep 14, 2026
Confidence3
Terminal-Bench 2.1
LMSpeed rank#4
Score86.6
UpdatedSep 14, 2026
Confidence3
SWE-bench ProV3 evidence
LMSpeed rank#5
Score67.7
UpdatedSep 14, 2026
Confidence3
NL2RepoV3 evidence
LMSpeed rank#3
Score55.9
UpdatedSep 14, 2026
Confidence3
FrontierSWE
LMSpeed rank#2
Score73.5
UpdatedSep 14, 2026
Confidence3
MLS-Bench Lite
LMSpeed rank#2
Score41.0
UpdatedSep 14, 2026
Confidence3
PaperBench
LMSpeed rank#1
Score93.0
UpdatedSep 14, 2026
Confidence3
QwenReactBench
LMSpeed rank#1
Score1724.0
UpdatedSep 14, 2026
Confidence3
VulcanBench v3
LMSpeed rank#10
Score81.2
UpdatedSep 14, 2026
Confidence3
OpenHarmony Bench
LMSpeed rank#1
Score60.8
UpdatedSep 14, 2026
Confidence3
FrontierSWE v2
LMSpeed rank#9
Score15.8
UpdatedSep 14, 2026
Confidence3
LiveCodeBench (Vals)
LMSpeed rank#9
Score87.9
UpdatedSep 14, 2026
Confidence3
SWE-bench (Vals)
LMSpeed rank#14
Score85.6
UpdatedSep 14, 2026
Confidence3
DeepSWE
LMSpeed rank#14
Score56.6
UpdatedSep 14, 2026
Confidence3
AA Coding Index
LMSpeed rank#17
Score71.8
UpdatedAug 14, 2026
Confidence1
AA-SciCodeV3 evidence
LMSpeed rank#27
Score52.9
UpdatedAug 14, 2026
Confidence1

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score63.580% interval53.3–73.62/4 Measured dimensionsGPQA92.7%#16 / 218HLE43.0%#18 / 216
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning62.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.79 · q 1.00
gpqa · gpqa · z 1.21 · q 1.00
hle · hle · z 0.94 · q 1.00
hle · hle_no_tools · z 0.10 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification64.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.05 · q 1.00
longbench · long_bench_v2 · z 2.81 · q 1.00
GPQAV3 evidence
LMSpeed rank#16
Score92.7%
UpdatedSep 14, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#18
Score43.0%
UpdatedSep 14, 2026
Confidence4
Reasoning score
LMSpeed rank#2
Score86.5
UpdatedSep 14, 2026
Confidence3
MRCRv2
LMSpeed rank#1
Score92.9
UpdatedSep 14, 2026
Confidence3
LongBench v2V3 evidence
LMSpeed rank#1
Score66.3
UpdatedSep 14, 2026
Confidence3
AA-LCRV3 evidence
LMSpeed rank#52
Score74.3
UpdatedAug 14, 2026
Confidence1
CritPtV3 evidence
LMSpeed rank#16
Score20.0
UpdatedAug 14, 2026
Confidence1

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score52.280% interval38.2–66.21/4 Measured dimensionsKnowledge score67.4#44 / 83GPQA-D92.6#11 / 32
Dimensions and evidence
Broad knowledge52.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z 0.23 · q 1.00
benchlm_category_knowledge · benchlm_category_knowledge · z -0.13 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge scoreV3 evidence
LMSpeed rank#44
Score67.4
UpdatedSep 14, 2026
Confidence3
GPQA-D
LMSpeed rank#11
Score92.6
UpdatedSep 14, 2026
Confidence3
HLE w/o tools
LMSpeed rank#7
Score43.6
UpdatedSep 14, 2026
Confidence3
GPQA Diamond (Vals)
LMSpeed rank#6
Score93.7
UpdatedSep 14, 2026
Confidence3
MMLU-Pro (Vals)
LMSpeed rank#15
Score88.6
UpdatedSep 14, 2026
Confidence3
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#2
Score58.1
UpdatedAug 14, 2026
Confidence1
AA-GPQA Diamond
LMSpeed rank#16
Score92.7
UpdatedAug 14, 2026
Confidence1
AA-HLE
LMSpeed rank#16
Score43.0
UpdatedAug 14, 2026
Confidence1
AA-Omniscience IndexV3 evidence
LMSpeed rank#34
Score3.4
UpdatedAug 14, 2026
Confidence1
AA-Omniscience Accuracy
LMSpeed rank#51
Score31.9
UpdatedAug 14, 2026
Confidence1
AA-Omniscience Hallucination Rate
LMSpeed rank#84
Score41.7
UpdatedAug 14, 2026
Confidence1

Math

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofsPrior only
Applied & tool-assisted mathPrior only

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Rated

Score64.3#180% interval58.0–70.64/4 Measured dimensionsMultimodal Grounded score87.4#6 / 56MMMU-Pro82.3#5 / 31
Dimensions and evidence
Perception & OCR54.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
simplevqa · simple_vqa · z 0.65 · q 1.00
Document & spatial understanding65.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
charxiv · charxiv · z 2.18 · q 1.00
Visual reasoning72.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
erqa · erqa · z 2.08 · q 1.00
mathvision · math_vision · z 3.00 · q 1.00
medxpert · med_xpert_qa_mm · z 0.74 · q 0.75
mmmu_pro · mmmu_pro · z 0.99 · q 1.00
Video & grounded action64.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
screenspot · screen_spot_pro · z 0.41 · q 1.00
video_mmmu · video_mmmu · z 3.00 · q 1.00
Multimodal Grounded score
LMSpeed rank#6
Score87.4
UpdatedSep 14, 2026
Confidence3
MMMU-ProV3 evidence
LMSpeed rank#5
Score82.3
UpdatedSep 14, 2026
Confidence3
MathVisionV3 evidence
LMSpeed rank#1
Score95.2
UpdatedSep 14, 2026
Confidence3
MathVision w/ Python
LMSpeed rank#2
Score97.7
UpdatedSep 14, 2026
Confidence3
BabyVision
LMSpeed rank#1
Score82.0
UpdatedSep 14, 2026
Confidence3
BabyVision w/ Python
LMSpeed rank#1
Score91.3
UpdatedSep 14, 2026
Confidence3
ZeroBench
LMSpeed rank#3
Score24.0
UpdatedSep 14, 2026
Confidence3
ZeroBench w/ Python
LMSpeed rank#1
Score49.0
UpdatedSep 14, 2026
Confidence3
MedXpertQA (MM)V3 evidence
LMSpeed rank#2
Score80.4
UpdatedSep 14, 2026
Confidence3
ScreenSpot ProV3 evidence
LMSpeed rank#4
Score84.5
UpdatedSep 14, 2026
Confidence3
Vision2Web
LMSpeed rank#1
Score69.0
UpdatedSep 14, 2026
Confidence3
CharXiv w/o tools
LMSpeed rank#1
Score88.4
UpdatedSep 14, 2026
Confidence3
CharXivV3 evidence
LMSpeed rank#1
Score93.5
UpdatedSep 14, 2026
Confidence3
OmniDocBench 1.5
LMSpeed rank#1
Score92.1
UpdatedSep 14, 2026
Confidence3
OCRBench V2
LMSpeed rank#1
Score74.2
UpdatedSep 14, 2026
Confidence3
CC-OCR
LMSpeed rank#3
Score79.6
UpdatedSep 14, 2026
Confidence3
RealWorldQA
LMSpeed rank#1
Score88.0
UpdatedSep 14, 2026
Confidence3
ERQAV3 evidence
LMSpeed rank#1
Score77.8
UpdatedSep 14, 2026
Confidence3
SimpleVQAV3 evidence
LMSpeed rank#3
Score75.0
UpdatedSep 14, 2026
Confidence3
PerceptionBench
LMSpeed rank#1
Score63.5
UpdatedSep 14, 2026
Confidence3
Video-MME (with subtitle)
LMSpeed rank#1
Score90.4
UpdatedSep 14, 2026
Confidence3
VideoMMMUV3 evidence
LMSpeed rank#1
Score88.7
UpdatedSep 14, 2026
Confidence3
MMVU
LMSpeed rank#1
Score82.4
UpdatedSep 14, 2026
Confidence3
MLVU (M-Avg)
LMSpeed rank#1
Score90.8
UpdatedSep 14, 2026
Confidence3
LVBench
LMSpeed rank#3
Score81.8
UpdatedSep 14, 2026
Confidence3
Design Arena Website
LMSpeed rank#17
Score1295.0
UpdatedSep 8, 2026
Confidence3
AA-MMMU-Pro
LMSpeed rank#9
Score82.3
UpdatedAug 14, 2026
Confidence1

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score57.980% interval41.9–73.91/4 Measured dimensionsInstruction Following score90.7#15 / 52IFBench82.8#1 / 14
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization57.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · if_bench · z 0.76 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#15
Score90.7
UpdatedSep 14, 2026
Confidence3
IFBenchV3 evidence
LMSpeed rank#1
Score82.8
UpdatedSep 14, 2026
Confidence3

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Alibaba
alibaba
$2/M$6/M100%undefined tokens / undefined tokens

Pricing Comparison

Compare Qwen3.8 Max API pricing across 117 providers. Prices range from $0.0001/request to $1027.40/M. FineOneAPI offers the lowest rate at $0.0001/request. 2 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
qwen3.8-max
qwen
$4.80/M
$14.40/M
L1
100%
qwen3.8-max
default
-59%$0.822/M
Cache read$0.103/M
-59%$2.47/M
L1
100%
qwen3.8-max
bailian
-57%$0.857/M
Cache read$0.107/M
-57%$2.57/M
L1
100%
qwen3.8-max
default
-18%$1.64/M
Cache read$0.205/MCache write$2.05/MCache write 1h$3.29/M
-18%$4.93/M
L1
100%
qwen3.8-max
default
-18%$1.64/M
Cache read$0.205/MCache write$2.05/MCache write 1h$3.29/M
-18%$4.93/M
L1
100%
qwen3.8-max
default
-59%$0.822/M
Cache read$0.103/M
-59%$2.47/M
L1
100%
qwen3.8-max
Self-Deployed-1
-91%$0.178/M
Cache read$0.022/M
-91%$0.534/M
L1
100%
qwen3.8-max
deepseek
-67%$0.658/M
-67%$1.97/M
L1
100%
L2
100%
qwen3.8-max
OpenModels
-40%$1.20/M
Cache read$0.012/M
-40%$3.60/M
L1
100%
[xj]q|abyss/qwen3.8-max
default
$6.00/request
-
L1
100%
[hm]q|满血/qwen3.8-max
default
$20.00/request
-
L1
100%
qwen3.8-max
default
-91%$0.171/M
Cache read$0.034/M
-91%$0.514/M
L1
99%
L2
100%
qwen3.8-max
default
Free
Free
L1
99%
L2
78%
Qwen-Ambassador/Qwen3.8-Max
diamond-glm
$54.75/M
$54.75/M
L1
100%
qwen3.8-max
default
$0.219/request
-
L1
100%
qwen3.8-max
default
-40%$1.20/M
Cache read$0.150/M
-40%$3.60/M
L1
100%
qwen3.8-max
default
$2.00/M
Cache read$0.222/M
-67%$2.00/M
L1
100%
qwen3.8-max
default
$2.16/M
$6.48/M
L1
100%
L2
0%
qwen3.8-max
default
$13.00/request
-
L1
100%
L2
100%
qwen3.8-max
default
-50%$1.00/M
-50%$3.00/M
Showing 20 model IDs of 58.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Qwen3.8 Max include?
LMSpeed shows Qwen3.8 Max benchmark context, API price, output speed, first-token latency, and provider data across 119 providers when those signals are available.
What is the Qwen3.8 Max API price?
Qwen3.8 Max has pricing from undefined provider} other undefined providers}}, ranging from $0.0001/request to $1027.40/M. FineOneAPI has the lowest listed price.
What does the Qwen3.8 Max API pricing table include?
The Qwen3.8 Max API pricing table compares 119 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Qwen3.8 Max API pricing?
FineOneAPI currently has the lowest listed Qwen3.8 Max price at $0.0001/request across undefined provider} other undefined providers}}.
Can I compare Qwen3.8 Max API price and speed together?
Yes. LMSpeed shows Qwen3.8 Max API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Qwen3.8 Max API free?
Yes, Qwen3.8 Max free API options are available through 2 providerundefined other undefined} on LMSpeed, including Zero API, DeadlySignal API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Qwen3.8 Max free API access?
LMSpeed currently lists 2 free API providerundefined other undefined} for Qwen3.8 Max: Zero API, DeadlySignal API. Check each provider row before using it because free tier limits can change.

Also known as

Qwen-Ambassador/Qwen3.8-MaxQwen3.8-MaxQwen3.8-Max-Preview[default]Qwen-Ambassador/Qwen3.8-Max[hm]q|满血/qwen3.8-max

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation