Qwen3.8 Max (0902) API Benchmarks, Pricing & Provider Data
Compare Qwen3.8 Max (0902) with another model
Choose a model to open its comparison page.
Qwen3.8 Max (0902) API pricing covers undefined API provider} other undefined API providers}}, from $0.178/M to $0.857/M.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Sep 2026
- Tokenizer
- Qwen
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Pricing Comparison
Compare Qwen3.8 Max (0902) API pricing across 7 providers. Prices range from $0.178/M to $0.857/M. Jeniya AI API offers the lowest rate at $0.178/M.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | qwen3.8-max-0902 | bailian | $0.857/M Cache read$0.107/M | $2.57/M | — | — | — | |
L1 100% | qwen3.8-max-0902 | Self-Deployed-1 | $0.178/M Cache read$0.022/M | $0.534/M | — | — | — | |
L1 100% | qwen3.8-max-0902 | Self-Deployed-1 | $0.178/M Cache read$0.022/M | $0.534/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Frequently Asked Questions
- What benchmark data does Qwen3.8 Max (0902) include?
- LMSpeed shows Qwen3.8 Max (0902) benchmark context, API price, output speed, first-token latency, and provider data across 7 providers when those signals are available.
- What is the Qwen3.8 Max (0902) API price?
- Qwen3.8 Max (0902) has pricing from undefined provider} other undefined providers}}, ranging from $0.178/M to $0.857/M. Jeniya AI API has the lowest listed price.
- What does the Qwen3.8 Max (0902) API pricing table include?
- The Qwen3.8 Max (0902) API pricing table compares 7 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Qwen3.8 Max (0902) API pricing?
- Jeniya AI API currently has the lowest listed Qwen3.8 Max (0902) price at $0.178/M across undefined provider} other undefined providers}}.
- Is Qwen3.8 Max (0902) API free?
- Qwen3.8 Max (0902) does not currently have a free API tier on LMSpeed. All 7 providers charge per token.
