Qwen3.8 Max API Benchmarks, Pricing & Provider Data
Compare Qwen3.8 Max with another model
Choose a model to open its comparison page.
Qwen3.8 Max benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0001/request. Qwen3.8 Max free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual ...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Multimodal64.3RatedGlobal rank #14/4 Measured dimensions
#2Reasoning63.5Estimated2/4 Measured dimensions
#3Coding61.7Estimated2/4 Measured dimensions
#4Agents61.5Estimated2/4 Measured dimensions
#5Instruction following57.9Provisional1/4 Measured dimensions
#6Knowledge52.2Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Aug 2026
- Tokenizer
- Qwen
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 14, 2026Overall
undefined metric} other undefined metrics}}
Overall score77.0#6 / 112SkillsBench70.2#1 / 3
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed43.0 tok/s#75 / 80Time to first token1.69 s#46 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$2.00/M#143 / 186Output price$6.00/M#128 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Estimated
Score61.580% interval50.6–72.42/4 Measured dimensionsAgentic score83.6#9 / 77Terminal-Bench 2.186.6#4 / 10
Agents
V3.0undefined metric} other undefined metrics}} · Estimated
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Score61.780% interval50.5–72.92/4 Measured dimensionsSciCode53.2%#23 / 89Coding score68.0#14 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score63.580% interval53.3–73.62/4 Measured dimensionsGPQA92.7%#16 / 218HLE43.0%#18 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score52.280% interval38.2–66.21/4 Measured dimensionsKnowledge score67.4#44 / 83GPQA-D92.6#11 / 32
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Rated
Score64.3#180% interval58.0–70.64/4 Measured dimensionsMultimodal Grounded score87.4#6 / 56MMMU-Pro82.3#5 / 31
Multimodal
V3.0undefined metric} other undefined metrics}} · Rated
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score57.980% interval41.9–73.91/4 Measured dimensionsInstruction Following score90.7#15 / 52IFBench82.8#1 / 14
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $2/M | $6/M | 100% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3.8 Max API pricing across 117 providers. Prices range from $0.0001/request to $1027.40/M. FineOneAPI offers the lowest rate at $0.0001/request. 2 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3.8-max | qwen | $4.80/M | $14.40/M | — | — | — | |
L1 100% | qwen3.8-max | default | -59%$0.822/M Cache read$0.103/M | -59%$2.47/M | — | — | — | |
L1 100% | qwen3.8-max | bailian | -57%$0.857/M Cache read$0.107/M | -57%$2.57/M | — | — | — | |
L1 100% | qwen3.8-max | default | -18%$1.64/M Cache read$0.205/MCache write$2.05/MCache write 1h$3.29/M | -18%$4.93/M | — | — | — | |
L1 100% | qwen3.8-max | default | -18%$1.64/M Cache read$0.205/MCache write$2.05/MCache write 1h$3.29/M | -18%$4.93/M | — | — | — | |
L1 100% | qwen3.8-max | default | -59%$0.822/M Cache read$0.103/M | -59%$2.47/M | — | — | — | |
L1 100% | qwen3.8-max | Self-Deployed-1 | -91%$0.178/M Cache read$0.022/M | -91%$0.534/M | — | — | — | |
L1 100% | qwen3.8-max | deepseek | -67%$0.658/M | -67%$1.97/M | — | — | — | |
L1 100% L2 100% | qwen3.8-max | OpenModels | -40%$1.20/M Cache read$0.012/M | -40%$3.60/M | — | — | — | |
L1 100% | [xj]q|abyss/qwen3.8-max | default | $6.00/request | - | — | — | — | |
L1 100% | [hm]q|满血/qwen3.8-max | default | $20.00/request | - | — | — | — | |
L1 100% | qwen3.8-max | default | -91%$0.171/M Cache read$0.034/M | -91%$0.514/M | — | — | — | |
DeadlySignal API Free | L1 99% L2 100% | qwen3.8-max | default | Free | Free | — | — | — |
L1 99% L2 78% | Qwen-Ambassador/Qwen3.8-Max | diamond-glm | $54.75/M | $54.75/M | — | — | — | |
L1 100% | qwen3.8-max | default | $0.219/request | - | — | — | ||
L1 100% | qwen3.8-max | default | -40%$1.20/M Cache read$0.150/M | -40%$3.60/M | — | — | — | |
L1 100% | qwen3.8-max | default | $2.00/M Cache read$0.222/M | -67%$2.00/M | — | — | — | |
L1 100% | qwen3.8-max | default | $2.16/M | $6.48/M | — | — | — | |
L1 100% L2 0% | qwen3.8-max | default | $13.00/request | - | — | — | ||
L1 100% L2 100% | qwen3.8-max | default | -50%$1.00/M | -50%$3.00/M | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.2
glm-5-2
Zhipu GLM-5.2 is Zhipu latest flagship coding and agentic model with a 1M-token context window, enhanced reasoning modes, and long-horizon software engineering capabilities.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Frequently Asked Questions
- What benchmark data does Qwen3.8 Max include?
- LMSpeed shows Qwen3.8 Max benchmark context, API price, output speed, first-token latency, and provider data across 119 providers when those signals are available.
- What is the Qwen3.8 Max API price?
- Qwen3.8 Max has pricing from undefined provider} other undefined providers}}, ranging from $0.0001/request to $1027.40/M. FineOneAPI has the lowest listed price.
- What does the Qwen3.8 Max API pricing table include?
- The Qwen3.8 Max API pricing table compares 119 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3.8 Max API pricing?
- FineOneAPI currently has the lowest listed Qwen3.8 Max price at $0.0001/request across undefined provider} other undefined providers}}.
- Can I compare Qwen3.8 Max API price and speed together?
- Yes. LMSpeed shows Qwen3.8 Max API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Qwen3.8 Max API free?
- Yes, Qwen3.8 Max free API options are available through 2 providerundefined other undefined} on LMSpeed, including Zero API, DeadlySignal API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3.8 Max free API access?
- LMSpeed currently lists 2 free API providerundefined other undefined} for Qwen3.8 Max: Zero API, DeadlySignal API. Check each provider row before using it because free tier limits can change.
