Qwen
·Released on 26 авг. 2026 г.

Qwen3.8 Flash API Benchmarks, Pricing & Provider Data

Compare Qwen3.8 Flash with another model

Choose a model to open its comparison page.

Share on X
LLM

Цены API Qwen3.8 Flash у поставщиков (36): от $0.000015/M до $100.00/request. Бесплатный API Qwen3.8 Flash предлагают поставщики: 2.

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart anal...

Стоимость
$0.000015/ 1M · 8:1 in:out
$0.0000024 in · $0.0000126 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1Mtokens
1.2K pages of text
OUTPUT
131.1Ktokens
8K128K1M4M
1M

Возможности

Technical Details

Вход
Выход
Выпущена
Aug 2026
Документация
Токенизатор
Qwen
Архитектура
text+image+video->text
Модерация
Нет
Поддерживаемые параметры
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Сравнение цен

Compare Qwen3.8 Flash API pricing across 34 providers. Prices range from $0.000015/M to $100.00/request. OAI2API offers the lowest rate at $0.000015/M. 2 providers offer free API credits or a free tier.

ПровайдерРаботоспособностьВариант моделиГруппаВходные данные ($/M)Выходные данные ($/M)Скорость (т/с)Первый токенАудит
L1
100%
qwen3.8-flash
bailian
$0.057/M
Cache read$0.0071/M
$0.193/M
L1
100%
qwen3.8-flash
qwen
$0.320/M
$1.08/M
L1
100%
L2
100%
qwen3.8-flash
Token
$0.150/M
Cache read$0.016/M
$0.470/M
L1
100%
L2
0%
qwen3.8-flash
default
$13.00/request
-
L1
100%
qwen3.8-flash
qwen-new
$0.099/M
Cache read$0.012/MCache write$0.154/MCache write 1h$0.247/M
$0.333/M
L1
100%
qwen3.8-flash
Qwen
$0.137/M
Cache read$0.014/M
$0.411/M
L1
100%
L2
100%
qwen3.8-flash
default
$0.800/M
Cache read$0.100/M
$2.70/M
L1
99%
qwen3.8-flash
gpt
$75.00/M
$75.00/M
L1
100%
L2
100%
qwen3.8-flash
default
$37.50/M
$37.50/M
L1
99%
L2
100%
qwen3.8-flash
diamond-glm
$54.75/M
$54.75/M
L1
99%
qwen3.8-flash
free
$0.000015/M
Cache write$0.0000016/MCache write 1h$0.00000256/M
$0.000047/M
L1
100%
qwen3.8-flash
Alibaba-2
$0.033/M
Cache read$0.0036/M
$0.105/M
L1
100%
qwen3.8-flash
default
$0.500/M
Cache read$0.050/MCache write$0.625/MCache write 1h$1.00/M
$1.50/M
L1
99%
qwen3.8-flash
国产
$30.00/M
$30.00/M
L1
100%
qwen3.8-flash
Alibaba-2
$0.033/M
Cache read$0.0036/M
$0.105/M
S3AI API
Бесплатно
L1
98%
L2
98%
qwen3.8-flash-free
free
Бесплатно
Бесплатно
L1
100%
qwen3.8-flash
default
$2.40/M
Cache read$0.300/M
$8.10/M
L1
100%
qwen3.8-flash
default
$38.00/M
$38.00/M
L1
99%
qwen3.8-flash
default
$4.00/M
Cache read$0.400/M
$20.00/M
初叶🍂Furry API
Бесплатно
L1
50%
qwen3.8-flash
free
Бесплатно
Бесплатно
Показано 20 ID моделей из 21.

Альтернативы и похожие модели

DeepSeek V4 Flash

deepseek-v4-flash

DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.

21 общих провайдеров

DeepSeek V4 Pro

deepseek-v4-pro

DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.

21 общих провайдеров

GPT-5.6 Luna

gpt-5-6-luna

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

21 общих провайдеров

GLM 5.3 Flash

glm-5-3-flash

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

21 общих провайдеров

GPT-5.4

gpt-5-4

OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.

20 общих провайдеров

Claude Opus 4.6

claude-opus-4-6

Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.

20 общих провайдеров

Frequently Asked Questions

What benchmark data does Qwen3.8 Flash include?
LMSpeed shows Qwen3.8 Flash benchmark context, API price, output speed, first-token latency, and provider data across 36 providers when those signals are available.
Сколько стоит API Qwen3.8 Flash?
Цены Qwen3.8 Flash у поставщиков (36): от $0.000015/M до $100.00/request. Самая низкая указанная цена — у OAI2API.
What does the Qwen3.8 Flash API pricing table include?
The Qwen3.8 Flash API pricing table compares 36 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
У какого провайдера самая низкая цена API Qwen3.8 Flash?
Сейчас самая низкая указанная цена Qwen3.8 Flash — $0.000015/M у OAI2API. Сравниваются поставщики: 36.
Является ли API Qwen3.8 Flash бесплатным?
Да, Qwen3.8 Flash доступна бесплатно через 2 API-провайдерundefined other undefined} на LMSpeed, включая 初叶🍂Furry API, S3AI API. Эти провайдеры предлагают бесплатные API-кредиты или бесплатный тариф без поперечной тарификации.
Where can I get Qwen3.8 Flash free API access?
LMSpeed currently lists 2 free API providerundefined other undefined} for Qwen3.8 Flash: 初叶🍂Furry API, S3AI API. Check each provider row before using it because free tier limits can change.

Также известна как

qwen3.8-flashqwen3.8-flash-freeqwen3.8‑flash

Рейтинги основаны на тестах, предоставленных сообществом, и периодических зондах работоспособности. Носит рекомендательный характер, не является официальными данными.Стандартные бенчмарки могут включать BenchLM и другие открытые источники.

Работайте с данными LMSpeed

Бесплатно

Бесплатный публичный API для цен LLM, бенчмарков и данных провайдеров

  • Актуальные цены и доступность
  • Бенчмарки скорости и задержки
  • 300+ моделей, 600+ провайдеров
  • 1 000 запросов в день
Документация API