DeepSeek
·Released on Apr 24, 2026

DeepSeek V4 Flash API Benchmarks, Pricing & Provider Data

Compare DeepSeek V4 Flash with another model

Choose a model to open its comparison page.

Share on X
LLM

DeepSeek V4 Flash benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0000044/M. DeepSeek V4 Flash free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reason...

Quality
#79of 112
51.0
LMSpeed score
Speed
#14of 80
85char/s
6.74 s
Cost
#18of 186
$0.0000044/ 1M · 8:1 in:out
$0.000000704 in · $0.0000037 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
5 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents45.4Coding43.9Reasoning53.3Knowledge41.4Math32.6Multilingual-Multimodal-Instruction following-
#1Reasoning53.3RatedGlobal rank #303/4 Measured dimensions
80% interval: 44.362.3
abstract logicPrior only
scientific causal59.8
gpqa · gpqa / hle · hle
multistep constraints54.2
mmlu_pro · mmlu_pro
evidence verification45.9
mrcr · mrcr1m
#2Agents45.4RatedGlobal rank #514/4 Measured dimensions
80% interval: 39.151.7
planning52.7
deep_planning · gert_labs
tool use45.3
mcp_atlas · mcp_atlas / toolathlon · toolathlon
environment execution37.9
browsecomp · browse_comp / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability45.7
claw_eval · claw_eval
#3Coding43.9RatedGlobal rank #374/4 Measured dimensions
80% interval: 37.350.5
code generation52.8
scicode · scicode
repository engineering38.9
swe_pro · swe_pro
debugging testing42
swe_multilingual · benchlm_coding_swe_multilingual / swe_verified · swe_verified
tooling quality42
terminalbench · benchlm_coding_terminal_bench2
#4Knowledge41.4Provisional1/4 Measured dimensions
80% interval: 23.459.5
broad knowledgePrior only
professional knowledgePrior only
factuality41.4
simpleqa · simple_qa
retrieval open bookPrior only
#5Math32.6Estimated2/4 Measured dimensions
80% interval: 20.444.8
foundational mathPrior only
competition math26.5
hmmt · hmmt_feb2026
proof frontier38.7
imo · imo_answer_bench
applied tool mathPrior only
No data:MultilingualMultimodalInstruction following

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1.0Mtokens
1.3K pages of text
OUTPUT
384Ktokens
8K128K1M4M
1.0M

Features

Technical Details

Input
Output
Released
Apr 2026
Tokenizer
DeepSeek
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_completion_tokensmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score51.0#79 / 112
Overall score
LMSpeed rank#79
Score51.0
UpdatedAug 13, 2026
Confidence2

Pricing

undefined metric} other undefined metrics}}

Input price$0.130/M#18 / 186Output price$0.280/M#10 / 186
Input price
LMSpeed rank#18
Score$0.130/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#10
Score$0.280/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score45.4#5180% interval39.1–51.74/4 Measured dimensionsAgentic score34.6#65 / 77Terminal-Bench 2.056.6#38 / 53
Dimensions and evidence
Planning & decomposition52.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.37 · q 1.00
Tool use45.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z -0.81 · q 1.00
toolathlon · toolathlon · z -0.75 · q 1.00
Environment & long-horizon execution37.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -1.88 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.89 · q 1.00
Recovery & completion reliability45.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
claw_eval · claw_eval · z -0.29 · q 1.00
Agentic score
LMSpeed rank#65
Score34.6
UpdatedAug 13, 2026
Confidence2
Terminal-Bench 2.0V3 evidence
LMSpeed rank#38
Score56.6
UpdatedAug 13, 2026
Confidence2
BrowseCompV3 evidence
LMSpeed rank#31
Score53.5
UpdatedAug 13, 2026
Confidence2
HLE w/ tools
LMSpeed rank#11
Score40.3
UpdatedAug 13, 2026
Confidence2
MCP AtlasV3 evidence
LMSpeed rank#21
Score67.4
UpdatedAug 13, 2026
Confidence2
ToolathlonV3 evidence
LMSpeed rank#16
Score43.5
UpdatedAug 13, 2026
Confidence2
Claw-EvalV3 evidence
LMSpeed rank#16
Score57.8
UpdatedAug 13, 2026
Confidence2
Gert LabsV3 evidence
LMSpeed rank#18
Score54.4
UpdatedAug 13, 2026
Confidence2

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score43.9#3780% interval37.3–50.54/4 Measured dimensionsSciCode50.3%#32 / 89Coding score46.0#65 / 87
Dimensions and evidence
Code generation52.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.20 · q 1.00
Repository engineering38.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_pro · swe_pro · z -1.38 · q 1.00
Debugging & testing42
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_multilingual · benchlm_coding_swe_multilingual · z -1.30 · q 1.00
swe_verified · swe_verified · z -0.59 · q 1.00
Tool-assisted development & quality42
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · benchlm_coding_terminal_bench2 · z -1.05 · q 1.00
SciCodeV3 evidence
LMSpeed rank#32
Score50.3%
UpdatedSep 13, 2026
Confidence4
Coding score
LMSpeed rank#65
Score46.0
UpdatedAug 13, 2026
Confidence2
Codeforces
LMSpeed rank#4
Score2816.0
UpdatedAug 13, 2026
Confidence2
SWE-bench VerifiedV3 evidence
LMSpeed rank#19
Score78.6
UpdatedAug 13, 2026
Confidence2
SWE-bench ProV3 evidence
LMSpeed rank#43
Score52.3
UpdatedAug 13, 2026
Confidence2
SWE MultilingualV3 evidence
LMSpeed rank#20
Score70.2
UpdatedAug 13, 2026
Confidence2
Terminal-Bench 2.0V3 evidence
LMSpeed rank#28
Score56.6
UpdatedAug 13, 2026
Confidence2
LiveCodeBench Pass@1-COT
LMSpeed rank#4
Score88.4
UpdatedAug 13, 2026
Confidence2

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score53.3#3080% interval44.3–62.33/4 Measured dimensionsMMLU-Pro86.4%#17 / 129GPQA86.7%#56 / 218
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning59.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
gpqa · gpqa · z 0.97 · q 1.00
hle · hle · z 0.83 · q 1.00
Multi-step constraints54.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 0.36 · q 1.00
Evidence integration & verification45.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mrcr · mrcr1m · z -0.45 · q 0.75
MMLU-ProV3 evidence
LMSpeed rank#17
Score86.4%
UpdatedAug 13, 2026
Confidence2
GPQAV3 evidence
LMSpeed rank#56
Score86.7%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#55
Score30.3%
UpdatedSep 13, 2026
Confidence4
MRCR 1MV3 evidence
LMSpeed rank#4
Score76.9
UpdatedAug 13, 2026
Confidence2
CorpusQA 1M
LMSpeed rank#3
Score59.3
UpdatedAug 13, 2026
Confidence2

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score41.480% interval23.4–59.51/4 Measured dimensionsKnowledge score52.8#69 / 83SimpleQA28.9#4 / 4
Dimensions and evidence
Broad knowledgePrior only
Professional knowledgePrior only
Factuality41.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
simpleqa · simple_qa · z -1.44 · q 0.17
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#69
Score52.8
UpdatedAug 13, 2026
Confidence2
SimpleQAV3 evidence
LMSpeed rank#4
Score28.9
UpdatedAug 13, 2026
Confidence2
Chinese-SimpleQA
LMSpeed rank#4
Score73.2
UpdatedAug 13, 2026
Confidence2
GPQA-D
LMSpeed rank#26
Score87.4
UpdatedAug 13, 2026
Confidence2

Math

V3.0

undefined metric} other undefined metrics}} · Estimated

Score32.680% interval20.4–44.82/4 Measured dimensionsHMMT Feb 202691.9#8 / 18IMOAnswerBench85.1#6 / 7
Dimensions and evidence
Foundational mathPrior only
Competition math26.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
hmmt · hmmt_feb2026 · z -3.00 · q 1.00
Advanced proofs38.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
imo · imo_answer_bench · z -2.00 · q 0.88
Applied & tool-assisted mathPrior only
HMMT Feb 2026V3 evidence
LMSpeed rank#8
Score91.9
UpdatedAug 13, 2026
Confidence2
IMOAnswerBenchV3 evidence
LMSpeed rank#6
Score85.1
UpdatedAug 13, 2026
Confidence2
Apex
LMSpeed rank#6
Score19.1
UpdatedAug 13, 2026
Confidence2
Apex Shortlist
LMSpeed rank#4
Score72.1
UpdatedAug 13, 2026
Confidence2
Math score
LMSpeed rank#11
Score78.4
UpdatedAug 13, 2026
Confidence2

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · No data

80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1230.0#43 / 78
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
Design Arena Website
LMSpeed rank#43
Score1230.0
UpdatedAug 13, 2026
Confidence2

Instruction following

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalizationPrior only
Multi-turn & long instructionsPrior only

OpenRouter endpoints

17 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Novita
novita/fp8
$0.140/M$0.280/M100.0%undefined tokens / undefined tokens
Wafer
wafer
$0.100/M$0.250/M100.0%undefined tokens / undefined tokens
Phala
phala
$0.200/M$0.400/M100.0%undefined tokens / undefined tokens
Baidu
baidu/fp8
$0.066/M$0.132/M100.0%undefined tokens / undefined tokens
AtlasCloud
atlas-cloud/fp4
$0.140/M$0.280/M99.9%undefined tokens / undefined tokens
NextBit
nextbit/fp8
$0.150/M$0.350/M99.9%undefined tokens / undefined tokens
DigitalOcean
digitalocean
$0.068/M$0.168/M99.9%undefined tokens / undefined tokens
Parasail
parasail/fp8
$0.140/M$0.280/M99.9%undefined tokens / undefined tokens
Alibaba
alibaba/fp8
$0.134/M$0.268/M99.8%undefined tokens / undefined tokens
GMICloud
gmicloud/fp8
$0.091/M$0.182/M99.8%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.130/M$0.280/M99.6%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.090/M$0.180/M99.4%undefined tokens / undefined tokens
Venice
venice
$0.097/M$0.193/M99.2%undefined tokens / undefined tokens
StreamLake
streamlake/fp8
$0.066/M$0.131/M98.8%undefined tokens / undefined tokens
Azure
azure/us
$0.210/M$0.560/M98.6%undefined tokens / undefined tokens
Mancer 2
mancer/fp8
$0.190/M$0.500/M97.6%undefined tokens / undefined tokens
OpenInference
open-inference/fp8
$0.050/M$0.120/M93.9%undefined tokens / undefined tokens

Pricing Comparison

Compare DeepSeek V4 Flash API pricing across 386 providers. Prices range from $0.0000044/M to $1071.43/M. AIGCBAR offers the lowest rate at $0.0000044/M. 22 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
deepseek-v4-flash
deepseek-sale
-45%$0.071/M
Cache read$0.0014/M
$0.286/M
122.0 t/s
2.34 s
L1
99%
deepseek-v4-flash
临时渠道
-89%$0.014/M
Cache read$0.0041/M
-85%$0.041/M
106.1 t/s
2.90 s
L1
99%
deepseek-ai/DeepSeek-V4-Flash
临时渠道
-89%$0.014/M
Cache read$0.0041/M
-85%$0.041/M
L1
99%
deepseek-ai/deepseek-v4-flash
临时渠道
-89%$0.014/M
Cache read$0.0041/M
-85%$0.041/M
L1
99%
deepseek-v4-flash-free
临时渠道2
-21%$0.103/M
-63%$0.103/M
L1
100%
deepseek-v4-flash
DeepSeek-OSS
-80%$0.026/M
Cache read$0.0008/M
-72%$0.078/M
105.8 t/s
15.63 s
L1
100%
deepseek-v4-flash-free
Free
Free
Free
L1
100%
L2
0%
deepseek-v4-flash-preview
Archived
$1.00/M
Cache read$0.020/M
$4.00/M
101.2 t/s
2.01 s
L1
100%
L2
100%
deepseek-v4-flash
Unlimited
Free
Free
95.3 t/s
8.10 s
L1
100%
L2
100%
deepseek-v4-flash-0731
Unlimited
Free
Free
L1
55%
deepseek-v4-flash
default
$75.00/M
$75.00/M
86.3 t/s
4.08 s
L1
55%
deepseek/deepseek-v4-flash-free
default
$75.00/M
$75.00/M
L1
100%
deepseek-v4-flash
default
$41.10/M
Cache read$1.37/MCache write$0/MCache write 1h$0/M
$123.29/M
84.8 t/s
6.39 s
L1
100%
deepseek-ai/DeepSeek-V4-Flash
default
$0.274/request
-
L1
100%
deepseek-ai/DeepSeek-V4-Flash-0731
default
$0.342/request
-
L1
100%
deepseek-v4-flash
default
$0.205/M
Cache read$0.0068/M
$0.616/M
73.8 t/s
2.82 s
L1
100%
deepseek-v4-flash
default
$0.132/M
Cache read$0.0042/M
$0.396/M
73.1 t/s
2.52 s
L1
100%
deepseek-v4-flash
default
$0.137/M
Cache read$0.014/M
-2%$0.274/M
62.1 t/s
2.19 s
L1
100%
L2
0%
deepseek-v4-flash
default
$1.00/M
Cache read$0.020/M
$4.00/M
56.4 t/s
3.93 s
L1
100%
L2
8%
deepseek/deepseek-v4-flash
default
-14%$0.112/M
Cache read$0.022/M
-20%$0.224/M
56.0 t/s
6.10 s
Showing 20 model IDs of 202.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does DeepSeek V4 Flash include?
LMSpeed shows DeepSeek V4 Flash benchmark context, API price, output speed, first-token latency, and provider data across 408 providers when those signals are available.
What is the DeepSeek V4 Flash API price?
DeepSeek V4 Flash has pricing from undefined provider} other undefined providers}}, ranging from $0.0000044/M to $1071.43/M. AIGCBAR has the lowest listed price.
What does the DeepSeek V4 Flash API pricing table include?
The DeepSeek V4 Flash API pricing table compares 408 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest DeepSeek V4 Flash API pricing?
AIGCBAR currently has the lowest listed DeepSeek V4 Flash price at $0.0000044/M across undefined provider} other undefined providers}}.
Can I compare DeepSeek V4 Flash API price and speed together?
Yes. LMSpeed shows DeepSeek V4 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is DeepSeek V4 Flash API free?
Yes, DeepSeek V4 Flash free API options are available through 22 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, DeadlySignal API, DeadlySignal API, DeadlySignal API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get DeepSeek V4 Flash free API access?
LMSpeed currently lists 22 free API providerundefined other undefined} for DeepSeek V4 Flash: 兔子API, 兔子API, DeadlySignal API, DeadlySignal API, DeadlySignal API. Check each provider row before using it because free tier limits can change.

Also known as

AMD/deepseek-v4-flashDeepSeek-V4-FlashDeepSeek-V4-Flash-0731DeepSeek-V4-Flash-2026-04-23DeepSeek-V4-Flash-fast

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation