Qwen
·Released on Sep 8, 2025

Qwen Plus 0728 API Benchmarks, Pricing & Provider Data

Compare Qwen Plus 0728 with another model

Choose a model to open its comparison page.

Share on X
LLM

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1Mtokens
1.2K pages of text
OUTPUT
32.8Ktokens
8K128K1M4M
1M

Features

Technical Details

Input
Output
Released
Sep 2025
Knowledge cutoff
2025-03-31
Documentation
Tokenizer
Qwen3
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltylogprobsmax_tokenspresence_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Alibaba
alibaba
$0.260/M$0.780/M100%undefined tokens / undefined tokens

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation