Qwen
·Released on Feb 16, 2026Qwen3.5 Plus 2026-02-15 API Benchmarks, Pricing & Provider Data
Compare Qwen3.5 Plus 2026-02-15 with another model
Choose a model to open its comparison page.
LLM
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference e...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
INPUT
1Mtokens
≈ 1.2K pages of text
OUTPUT
65.5Ktokens
8K128K1M4M
1M
Features
Technical Details
- Input
- Output
- Released
- Feb 2026
- Tokenizer
- Qwen3
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.260/M | $1.56/M | 100% | — | — | undefined tokens / undefined tokens |
