Choose a model to open its comparison page.
Qwen3 VL 8B Thinking API pricing covers 23 API providers, from $0.027/M to $0.191/M.
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Qwen3 VL 8B Thinking API pricing across 23 providers. Prices range from $0.027/M to $0.191/M. Zhongzhuan Chat offers the lowest rate at $0.027/M.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3-vl-8b-thinking | default | $0.064/M | $0.637/M | — | — | — | |
L1 100% | qwen3-vl-8b-thinking | default | $0.064/M | $0.637/M | — | — | — | |
L1 100% | qwen3-vl-8b-thinking | default | $0.137/M | $1.37/M | — | — | — | |
L1 100% | qwen3-vl-8b-thinking | default | $0.034/M | $0.342/M | — | — | — |
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.