Qwen
·Released on Feb 25, 2026Qwen3.5-Flash API Benchmarks, Pricing & Provider Data
Compare Qwen3.5-Flash with another model
Choose a model to open its comparison page.
LLM
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference effic...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
INPUT
1Mtokens
≈ 1.2K pages of text
OUTPUT
65.5Ktokens
8K128K1M4M
1M
Features
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.065/M | $0.260/M | 100.0% | — | — | undefined tokens / undefined tokens |
