Qwen3.5 Max API Benchmarks, Pricing & Provider Data
Compare Qwen3.5 Max with another model
Choose a model to open its comparison page.
Qwen3.5 Max API pricing covers 5 API providers, from $0.500/M to $150.00/M. The page also shows measured API speed and first-token latency.
Alibaba Qwen3.5 Max is a high-capability language model in the Qwen series, offering enhanced reasoning, code generation, and multimodal capabilities.
Specifications
Pricing Comparison
Compare Qwen3.5 Max API pricing across 5 providers. Prices range from $0.500/M to $150.00/M. GLM BigModel Relay offers the lowest rate at $0.500/M.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 99% | qwen/qwen3.5-max | other | $0.500/M | $2.50/M | — | — | — | |
L1 100% | qwen3.5-max-2026-03-08 | default | $75.00/M | $75.00/M | — | — | — | |
L1 0% | qwen3.5-max-2026-03-08 | 2api | $1.00/request | - | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Frequently Asked Questions
- What benchmark data does Qwen3.5 Max include?
- LMSpeed shows Qwen3.5 Max benchmark context, API price, output speed, first-token latency, and provider data across 5 providers when those signals are available.
- What is the Qwen3.5 Max API price?
- Qwen3.5 Max has pricing from 5 providers, ranging from $0.500/M to $150.00/M. GLM BigModel Relay has the lowest listed price.
- What does the Qwen3.5 Max API pricing table include?
- The Qwen3.5 Max API pricing table compares 5 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3.5 Max API pricing?
- GLM BigModel Relay currently has the lowest listed Qwen3.5 Max price at $0.500/M across 5 providers.
- Can I compare Qwen3.5 Max API price and speed together?
- Yes. LMSpeed shows Qwen3.5 Max API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Qwen3.5 Max API free?
- Qwen3.5 Max does not currently have a free API tier on LMSpeed. All 5 providers charge per token.
