Qwen
·Released on Sep 8, 2025Qwen Plus 0728 API Benchmarks, Pricing & Provider Data
Compare Qwen Plus 0728 with another model
Choose a model to open its comparison page.
LLM
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
INPUT
1Mtokens
≈ 1.2K pages of text
OUTPUT
32.8Ktokens
8K128K1M4M
1M
Features
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba | $0.260/M | $0.780/M | 100% | — | — | undefined tokens / undefined tokens |
