Qwen3.6 Flash API Benchmarks, Pricing & Provider Data
Compare Qwen3.6 Flash with another model
Choose a model to open its comparison page.
Qwen3.6 Flash API pricing covers undefined API provider} other undefined API providers}}, from $0.013/M to $300.00/M. Qwen3.6 Flash free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Apr 2026
- Tokenizer
- Qwen3
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.188/M | $1.13/M | 100% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Qwen3.6 Flash API pricing across 45 providers. Prices range from $0.013/M to $300.00/M. Zero API offers the lowest rate at $0.013/M. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3.6-flash | default | $0.131/M | $0.789/M | — | — | — | |
L1 100% | qwen3.6-flash | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.6-flash-2026-04-16 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.6-flash | default | $0.013/M Cache read$0.0013/M | $0.077/M | — | — | — | |
L1 100% | qwen3.6-flash | 🇨🇳国产官方模型 | $0.720/M | $4.32/M | — | — | — | |
L1 100% | qwen3.6-flash | ALIYUN | $0.103/M | $0.619/M | — | — | — | |
L1 99% L2 100% | qwen3.6-flash | diamond-glm | $0.137/M Cache read$0.014/MCache write$0.171/MCache write 1h$0.274/M | $0.821/M | — | — | — | |
L1 100% | qwen3.6-flash | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.6-flash-2026-04-16 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3.6-flash-2026-04-16 | default | $37.50/M | $37.50/M | — | — | — | |
L1 100% | qwen3.6-flash | default | $37.50/M | $37.50/M | — | — | — | |
L1 99% | qwen3.6-flash | default | $60.00/M | $60.00/M | — | — | — | |
L1 100% | qwen3.6-flash | default | $71.92/M | $71.92/M | — | — | — | |
L1 100% | qwen3.6-flash | 阿里云Qwen | $48.75/M | $48.75/M | — | — | — | |
L1 99% | qwen3.6-flash | 国模分组 | $10.27/M | $10.27/M | — | — | — | |
L1 99% | qwen3.6-flash | default | $4.00/M Cache read$0.400/M | $20.00/M | — | — | — | |
L1 100% | qwen3.6-flash | vip | $1.20/M | $7.20/M | — | — | — | |
L1 0% | qwen3.6-flash | default | $0.132/M Cache read$0.013/M | $0.793/M | — | — | — | |
L1 0% | qwen3.6-flash | default | $0.188/M Cache read$0.019/M | $1.13/M | — | — | — |
Alternatives & Similar Models
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Qwen3.6 Plus
qwen3-6-plus
Alibaba Qwen3.6 Plus is an enhanced Qwen3.6-tier model optimized for reasoning, coding, and long-context tasks with balanced cost and performance.
Kimi K2.6
kimi-k2-6
Moonshot Kimi K2.6 is an open-source native multimodal agent model with 1 trillion MoE parameters, a 256K context window, and state-of-the-art coding, vision, and long-horizon agent swarm capabilities.
Frequently Asked Questions
- What benchmark data does Qwen3.6 Flash include?
- LMSpeed shows Qwen3.6 Flash benchmark context, API price, output speed, first-token latency, and provider data across 46 providers when those signals are available.
- What is the Qwen3.6 Flash API price?
- Qwen3.6 Flash has pricing from undefined provider} other undefined providers}}, ranging from $0.013/M to $300.00/M. Zero API has the lowest listed price.
- What does the Qwen3.6 Flash API pricing table include?
- The Qwen3.6 Flash API pricing table compares 46 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3.6 Flash API pricing?
- Zero API currently has the lowest listed Qwen3.6 Flash price at $0.013/M across undefined provider} other undefined providers}}.
- Can I compare Qwen3.6 Flash API price and speed together?
- Yes. LMSpeed shows Qwen3.6 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Qwen3.6 Flash API free?
- Yes, Qwen3.6 Flash free API options are available through 1 providerundefined other undefined} on LMSpeed, including Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Qwen3.6 Flash free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Qwen3.6 Flash: Zero API. Check each provider row before using it because free tier limits can change.
