Choose a model to open its comparison page.
Qwen3 Coder API pricing covers 110 API providers, from $0.0041/M to $19980.00/M. Qwen3 Coder free API options are available from 5 providers. The page also shows measured API speed and first-token latency.
Alibaba Qwen3 Coder is a code-specialized variant in the Qwen series, optimized for code generation, debugging, and software development tasks.
Input and output token limits for this model, plus how it ranks on long-context understanding.
Сторонние данные эндпоинтов OpenRouter показаны отдельно от измерений LMSpeed. Некоторые 30-минутные live-метрики появляются только после синхронизации с OpenRouter API key.
| Provider endpoint | Вход | Вывод | Доступность 1 день | Задержка 30м | Пропускная способность 30м | Контекст / вывод |
|---|---|---|---|---|---|---|
Alibaba alibaba/opensource | $0.975/M | $4.88/M | 100.0% | — | — | 262.1K токенов / 65.5K токенов |
Novita novita/fp8 | $0.380/M | $1.55/M | 99.8% | — | — | 262.1K токенов / 65.5K токенов |
DeepInfra deepinfra/turbo | $0.300/M | $1/M | 99.3% | — | — | 262.1K токенов / 65.5K токенов |
WandB wandb/bf16 | $1/M | $1.50/M | 98.4% | — | — | 262.1K токенов / 262.1K токенов |
Venice venice/fp8 | $0.350/M | $1.50/M | 96.0% | — | — | 256K токенов / 65.5K токенов |
Google google-vertex/us-south1 | $0.220/M | $1.80/M | 57.7% | — | — | 262.1K токенов / 65.5K токенов |
Compare Qwen3 Coder API pricing across 105 providers. Prices range from $0.0041/M to $19980.00/M. 6345ywz API offers the lowest rate at $0.0041/M. 5 providers offer free API credits or a free tier.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3-coder | default | $0.411/M | $1.64/M | — | — | — | |
L1 100% | qwen/qwen3-coder:free | default | $0.0050/request | - | — | — | — | |
L1 100% | qwen3-coder | default | $0.822/M | $3.29/M | — | — | — | |
L1 100% | qwen3-coder | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen3-coder | default | $0.411/M | $1.64/M | — | — | — | |
L1 100% | qwen3-coder | default | $0.764/M | $3.06/M | — | — | — | |
L1 100% | qwen3-coder | default | $0.764/M | $3.06/M | — | — | — | |
L1 100% | qwen/qwen3-coder | default | $1.96/M | $7.83/M | — | — | — | |
L1 100% | qwen3-coder | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | qwen/qwen3-coder | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | qwen3-coder | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | qwen/qwen3-coder:free | default | $75.00/M | $75.00/M | — | — | — | |
Dext API Бесплатно | L1 100% | qwen3-coder | 公益 | Бесплатно | Бесплатно | — | — | — |
L1 100% | qwen3-coder | default | $0.658/M | $2.63/M | — | — | — | |
L1 100% | qwen3-coder-480b | deepseek | $0.877/M | $1.75/M | — | — | — | |
L1 99% | qwen/qwen3-coder | 0倍倍率分组 | $0.0041/M Cache read$0.0004/M | $0.016/M | — | — | — | |
L1 100% | qwen3-coder | default | $3.60/M | $14.40/M | — | — | — | |
L1 100% | qwen3-coder:480b | default | $6.00/request | - | — | — | — | |
L1 98% | qwen/qwen3-coder | other | $0.125/M Cache read$0.012/M | $0.500/M | — | — | — | |
L1 98% | qwen3-coder | other | $0.187/M | $0.750/M | — | — | — |
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.