Sponsored byFusecodeEnterprise coding API for Claude Code, Codex, and model workflows.
LogoLMSpeed
  • Free
  • Models
  • Providers
  • Leaderboard
LogoLMSpeed
  1. Home
  2. Models
  3. Deepseek V4 Flash
LogoLMSpeed

The best API speed test tool

GitHubGitHubTwitterX (Twitter)Email
Product
  • Features
  • Pricing
  • FAQ
Leaderboard
  • Overview
  • Speed Ranking
  • Latency Ranking
  • Health Ranking
  • Model Pricing
  • Model Speed
  • Reasoning
  • Coding
Models
  • All Models
  • GPT
  • Claude
  • Gemini
  • DeepSeek
  • Llama
  • Qwen
Free Models
  • All Free Models
  • Free GPT
  • Free Claude
  • Free Gemini
  • Free DeepSeek
  • Free Llama
  • Free Qwen
Tools
  • Speed Test
  • Provider Audit
Company
  • About
Resources
  • Provider Directory
  • Documentation
  • Public API
  • Botab
  • VidBee
Legal
  • Cookie Policy
  • Privacy Policy
  • Terms of Service
© 2026 LMSpeed All Rights Reserved.Made by Nexmoe with ❤️
DeepSeek
DeepSeek
·Released on Apr 24, 2026

DeepSeek V4 Flash API Benchmarks, Pricing & Provider Data

Compare DeepSeek V4 Flash with another model

Choose a model to open its comparison page.

X (Twitter)Share on X
LLM

DeepSeek V4 Flash benchmark, API pricing, and provider data cover 381 API providers, with prices starting at $0.0000014/M. DeepSeek V4 Flash free API options are available from 9 providers. The page also shows measured API speed and first-token latency.

DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reason...

Quality
#47of 93
56.0
LMSpeed score
Speed
#36of 75
73char/s
6.51 s
Cost
#18of 174
$0.0000014/ 1M · 8:1 in:out
$0.000000224 in · $0.00000118 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
5 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents50.3Coding50.7Reasoning56.1Knowledge39.5Math35.5Multilingual-Multimodal-Instruction following-
#1Reasoning56.1Rated·Global rank #14·3/4 Measured dimensions
80% interval: 47.8–64.5
abstract logicPrior only
scientific causal62.1
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraints54.4
mmlu_pro · mmlu_pro
evidence verification51.8
lcr · lcr / mrcr · mrcr1m
#2Coding50.7Rated·Global rank #26·4/4 Measured dimensions
80% interval: 44.4–57.0
code generation60
scicode · scicode
repository engineering53.1
nl2repo · nl2_repo / swe_pro · swe_pro
debugging testing44
swe_multilingual · benchlm_coding_swe_multilingual / swe_verified · swe_verified
tooling quality45.7
terminalbench · aa_terminal_bench21 / terminalbench · benchlm_coding_terminal_bench2
#3Agents50.3Rated·Global rank #33·4/4 Measured dimensions
80% interval: 44.6–56.0
planning52.8
deep_planning · gert_labs
tool use52.8
mcp_atlas · mcp_atlas / tau · aa_tau3_banking / toolathlon · toolathlon
environment execution43.8
browsecomp · browse_comp / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability51.9
claw_eval · claw_eval / cyber_gym · cyber_gym
#4Knowledge39.5Provisional·1/4 Measured dimensions
80% interval: 25.4–53.5
broad knowledge39.5
aa_omniscience · aa_omniscience_index / benchlm_category_knowledge · benchlm_category_knowledge
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
#5Math35.5Estimated·2/4 Measured dimensions
80% interval: 23.2–47.8
foundational mathPrior only
competition math27.1
hmmt · hmmt_feb2026
proof frontier43.9
imo · imo_answer_bench
applied tool mathPrior only
No data:MultilingualMultimodalInstruction following
Category leaderboardsAgentsCodingReasoningKnowledgeMath

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1.0Mtokens
≈ 1.3K pages of text
OUTPUT
393.2Ktokens
8K128K1M4M
1.0M

Features

Technical Details

Input
Output
Released
Apr 2026
Documentation
OpenRouterHuggingFaceOllama
Tokenizer
DeepSeek
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Aug 10, 2026

Overall

2 metrics

Overall score56.0#47 / 93DeepSWE54.4#8 / 12
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
Overall score
LMSpeed rank#47
Score56.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
DeepSWE
LMSpeed rank#8
Score54.4
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Pricing

2 metrics

Input price$0.140/M#18 / 174Output price$0.280/M#8 / 174
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
Input price
LMSpeed rank#18
Score$0.140/M
Sourceartificialanalysis.ai
UpdatedAug 10, 2026
Confidence4
Output price
LMSpeed rank#8
Score$0.280/M
Sourceartificialanalysis.ai
UpdatedAug 10, 2026
Confidence4

Agents

V3.0

17 metrics · Rated

Score50.3#3380% interval44.6–56.04/4 Measured dimensionsAgentic score46.9#46 / 62Terminal-Bench 2.056.9#36 / 51
Dimensions and evidence
Planning & decomposition52.8
1 family · 1 metric
deep_planning · gert_labs · z 0.37 · q 1.00
Tool use52.8
3 families · 3 metrics
mcp_atlas · mcp_atlas · z -0.58 · q 1.00
tau · aa_tau3_banking · z 0.87 · q 1.00
toolathlon · toolathlon · z -0.62 · q 1.00
Environment & long-horizon execution43.8
3 families · 3 metrics
browsecomp · browse_comp · z -1.66 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z 0.09 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.85 · q 1.00
Recovery & completion reliability51.9
2 families · 2 metrics
claw_eval · claw_eval · z -0.33 · q 1.00
cyber_gym · cyber_gym · z 0.37 · q 1.00
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
Agentic score
LMSpeed rank#46
Score46.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Terminal-Bench 2.0V3 evidence
LMSpeed rank#36
Score56.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
BrowseCompV3 evidence
LMSpeed rank#20
Score73.2
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
HLE w/ tools
LMSpeed rank#7
Score45.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
MCP AtlasV3 evidence
LMSpeed rank#18
Score69.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
GDPval-AAV3 evidence
LMSpeed rank#27
Score1189.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
ToolathlonV3 evidence
LMSpeed rank#13
Score47.8
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA Agentic Index
LMSpeed rank#9
Score45.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
GDPval-AA
LMSpeed rank#9
Score52.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Claw-EvalV3 evidence
LMSpeed rank#17
Score57.8
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Gert LabsV3 evidence
LMSpeed rank#18
Score54.4
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA Tau3 BankingV3 evidence
LMSpeed rank#5
Score31.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Terminal-Bench 2.1
LMSpeed rank#3
Score82.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
CyberGymV3 evidence
LMSpeed rank#6
Score76.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Toolathlon-Verified
LMSpeed rank#4
Score70.3
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Agents' Last Exam
LMSpeed rank#2
Score25.2
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AutomationBench
LMSpeed rank#4
Score25.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Coding

V3.0

15 metrics · Rated

Score50.7#2680% interval44.4–57.04/4 Measured dimensionsSciCode42.0%#59 / 197Coding score46.7#53 / 74
Dimensions and evidence
Code generation60
1 family · 1 metric
scicode · scicode · z 1.17 · q 1.00
Repository engineering53.1
2 families · 2 metrics
nl2repo · nl2_repo · z 2.22 · q 1.00
swe_pro · swe_pro · z -1.47 · q 1.00
Debugging & testing44
2 families · 2 metrics
swe_multilingual · benchlm_coding_swe_multilingual · z -0.91 · q 1.00
swe_verified · swe_verified · z -0.57 · q 1.00
Tool-assisted development & quality45.7
1 family · 2 metrics
terminalbench · aa_terminal_bench21 · z -0.32 · q 1.00
terminalbench · benchlm_coding_terminal_bench2 · z -1.07 · q 1.00
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
SciCodeV3 evidence
LMSpeed rank#59
Score42.0%
Sourceartificialanalysis.ai
UpdatedAug 10, 2026
Confidence4
Coding score
LMSpeed rank#53
Score46.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Codeforces
LMSpeed rank#2
Score3052.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
SWE-bench VerifiedV3 evidence
LMSpeed rank#16
Score79.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
SWE-bench ProV3 evidence
LMSpeed rank#39
Score52.6
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
SWE MultilingualV3 evidence
LMSpeed rank#13
Score73.3
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Terminal-Bench 2.0V3 evidence
LMSpeed rank#26
Score56.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA Coding Index
LMSpeed rank#16
Score69.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA-SciCodeV3 evidence
LMSpeed rank#26
Score49.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
LiveCodeBench Pass@1-COT
LMSpeed rank#2
Score91.6
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Terminal-Bench 2.1
LMSpeed rank#3
Score82.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
NL2RepoV3 evidence
LMSpeed rank#2
Score54.2
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
DSBench-FullStack
LMSpeed rank#1
Score68.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
DSBench-Hard
LMSpeed rank#1
Score59.6
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA Terminal-Bench 2.1V3 evidence
LMSpeed rank#10
Score78.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Reasoning

V3.0

7 metrics · Rated

Score56.1#1480% interval47.8–64.53/4 Measured dimensionsMMLU-Pro86.2%#16 / 125GPQA86.7%#43 / 200
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning62.1
3 families · 3 metrics
critpt · critpt · z 0.96 · q 1.00
gpqa · gpqa · z 1.06 · q 1.00
hle · hle · z 1.00 · q 1.00
Multi-step constraints54.4
1 family · 1 metric
mmlu_pro · mmlu_pro · z 0.38 · q 1.00
Evidence integration & verification51.8
2 families · 2 metrics
lcr · lcr · z 0.05 · q 1.00
mrcr · mrcr1m · z -0.39 · q 0.50
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
MMLU-ProV3 evidence
LMSpeed rank#16
Score86.2%
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
GPQAV3 evidence
LMSpeed rank#43
Score86.7%
Sourceartificialanalysis.ai
UpdatedAug 10, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#42
Score30.3%
Sourceartificialanalysis.ai
UpdatedAug 10, 2026
Confidence4
MRCR 1MV3 evidence
LMSpeed rank#2
Score78.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
CorpusQA 1M
LMSpeed rank#2
Score60.5
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA-LCRV3 evidence
LMSpeed rank#45
Score65.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
CritPtV3 evidence
LMSpeed rank#16
Score16.6
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Knowledge

V3.0

9 metrics · Provisional

Score39.580% interval25.4–53.51/4 Measured dimensionsKnowledge score60.1#48 / 64SimpleQA34.1#2 / 2
Dimensions and evidence
Broad knowledge39.5
2 families · 2 metrics
aa_omniscience · aa_omniscience_index · z -0.26 · q 1.00
benchlm_category_knowledge · benchlm_category_knowledge · z -2.27 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
Knowledge scoreV3 evidence
LMSpeed rank#48
Score60.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
SimpleQA
LMSpeed rank#2
Score34.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Chinese-SimpleQA
LMSpeed rank#2
Score78.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
GPQA-D
LMSpeed rank#20
Score88.1
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#17
Score49.9
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA-GPQA Diamond
LMSpeed rank#20
Score90.8
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA-HLE
LMSpeed rank#22
Score36.8
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA-Omniscience Accuracy
LMSpeed rank#37
Score37.2
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
AA-Omniscience Hallucination Rate
LMSpeed rank#26
Score84.4
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Math

V3.0

5 metrics · Estimated

Score35.580% interval23.2–47.82/4 Measured dimensionsHMMT Feb 202694.8#3 / 16IMOAnswerBench88.4#3 / 5
Dimensions and evidence
Foundational mathPrior only
Competition math27.1
1 family · 1 metric
hmmt · hmmt_feb2026 · z -3.00 · q 1.00
Advanced proofs43.9
1 family · 1 metric
imo · imo_answer_bench · z -1.24 · q 0.63
Applied & tool-assisted mathPrior only
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
HMMT Feb 2026V3 evidence
LMSpeed rank#3
Score94.8
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
IMOAnswerBenchV3 evidence
LMSpeed rank#3
Score88.4
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Apex
LMSpeed rank#3
Score33.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Apex Shortlist
LMSpeed rank#2
Score85.7
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2
Math score
LMSpeed rank#6
Score81.6
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Multilingual

V3.0

0 metrics · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

1 metric · No data

80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1231.0#37 / 70
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
Benchmark
LMSpeed rank
Score
Source
Updated
Confidence
Design Arena Website
LMSpeed rank#37
Score1231.0
Sourcebenchlm.ai
UpdatedAug 10, 2026
Confidence2

Instruction following

V3.0

0 metrics · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalizationPrior only
Multi-turn & long instructionsPrior only

OpenRouter endpoints

20 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Cloudflare
cloudflare
$0.140/M$0.280/M100.0%——384K tokens / 384K tokens
DeepInfra
deepinfra/fp4
$0.090/M$0.180/M99.7%——1.0M tokens / 65.5K tokens
Baidu
baidu/fp8
$0.069/M$0.137/M99.6%——1.0M tokens / 131.1K tokens
DeepSeek
deepseek
$0.140/M$0.280/M99.6%——1.0M tokens / 384K tokens
DigitalOcean
digitalocean
$0.068/M$0.168/M99.4%——1.0M tokens / —
GMICloud
gmicloud/fp8
$0.094/M$0.188/M99.4%——1.0M tokens / —
Venice
venice
$0.138/M$0.275/M99.4%——1M tokens / 32.8K tokens
Novita
novita/fp8
$0.140/M$0.280/M99.3%——1.0M tokens / 393.2K tokens
AtlasCloud
atlas-cloud/fp4
$0.140/M$0.280/M99.0%——1.0M tokens / 393.2K tokens
StreamLake
streamlake/fp8
$0.068/M$0.137/M98.9%——1.0M tokens / 384K tokens
CoreWeave
coreweave/fp8
$0.140/M$0.280/M98.7%——1.0M tokens / 1.0M tokens
Alibaba
alibaba/fp8
$0.134/M$0.268/M98.7%——1M tokens / 393.2K tokens
Parasail
parasail/fp8
$0.140/M$0.280/M98.3%——1.0M tokens / 1.0M tokens
SiliconFlow
siliconflow/fp8
$0.130/M$0.280/M97.9%——1.0M tokens / 393.2K tokens
Fireworks
fireworks
$0.140/M$0.280/M97.7%——1.0M tokens / —
Morph
morph
$0.139/M$0.278/M95.8%——1.0M tokens / 1.0M tokens
Mancer 2
mancer/fp4
$0.175/M$0.500/M92.4%——1.0M tokens / 1.0M tokens
OpenInference
open-inference/fp8
$0.070/M$0.180/M92.0%——1.0M tokens / 393.2K tokens
Phala
phala
$0.200/M$0.400/M91.8%——1.0M tokens / 393.2K tokens
Ambient
ambient/fp4
$0.140/M$0.280/M0%——1.0M tokens / 1.0M tokens

Pricing Comparison

Compare DeepSeek V4 Flash API pricing across 372 providers. Prices range from $0.0000014/M to $1398.60/M. AIGCBAR offers the lowest rate at $0.0000014/M. 9 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
6345ywz API
L1
97%
DeepSeekdeepseek-v4-flash
临时渠道
-90%$0.014/M
Cache read$0.0041/M
-85%$0.041/M
106.1 t/s
2.90 s
—
L1
97%
DeepSeekdeepseek-ai/deepseek-v4-flash
临时渠道
-90%$0.014/M
Cache read$0.0041/M
-85%$0.041/M
—
—
—
L1
97%
DeepSeekdeepseek-ai/DeepSeek-V4-Flash
临时渠道
-90%$0.014/M
Cache read$0.0041/M
-85%$0.041/M
—
—
—
Fengsili API
L1
59%
DeepSeekdeepseek-v4-flash
default
$75.00/M
$75.00/M
86.3 t/s
4.08 s
—
L1
59%
DeepSeekdeepseek-ai/DeepSeek-V4-Flash
default
$75.00/M
$75.00/M
—
—
—
钠 API
L1
100%
DeepSeekdeepseek-v4-flash
default
-2%$0.137/M
Cache read$0.0027/MCache write$0/MCache write 1h$0/M
-2%$0.274/M
84.8 t/s
6.39 s
—
L1
100%
DeepSeekdeepseek-ai/DeepSeek-V4-Flash
default
$0.0027/request
-
—
—
—
YUNWU API
L1
100%
DeepSeekdeepseek-v4-flash
default
-51%$0.068/M
Cache read$0.0014/M
-51%$0.137/M
73.8 t/s
2.82 s
846886100
Tokeness.io
L1
100%
DeepSeekdeepseek-v4-flash
default
-43%$0.080/M
Cache read$0.0017/M
-43%$0.160/M
73.1 t/s
2.52 s
10084100100
PICO API
L1
100%
DeepSeekdeepseek-v4-flash
default
-2%$0.137/M
Cache read$0.0027/M
-2%$0.274/M
62.1 t/s
2.19 s
10068100100
ChooseC API
L1
99%
DeepSeekdeepseek-v4-flash
default
$1.00/M
Cache read$0.020/M
$2.00/M
56.4 t/s
3.93 s
—
向量引擎
L1
100%
DeepSeekdeepseek-v4-flash
default
-51%$0.068/M
Cache read$0.0014/M
-51%$0.137/M
53.6 t/s
3.26 s
1008484100
L1
100%
DeepSeekdeepseek-v4-flash-202605
测试
$0.685/M
Cache read$0.014/M
$1.37/M
—
—
—
天絮 API
L1
99%
DeepSeekdeepseek-v4-flash
default
$0.950/M
Cache read$0.020/M
$1.90/M
49.0 t/s
2.64 s
—
L1
99%
DeepSeekdeepseek-ai/DeepSeek-V4-Flash
default
$0.643/M
Cache read$0.129/M
$1.29/M
—
—
—
9527 API
L1
100%
DeepSeekdeepseek-v4-flash
free
-29%$0.100/M
Cache read$0.0020/M
-29%$0.200/M
31.1 t/s
7.88 s
—
小老鼠的奶酪工坊-酒馆聊天api
L1
99%
DeepSeekdeepseek-v4-flash
鱼干公益基础组
-93%$0.010/M
Cache read$0.0050/M
-96%$0.010/M
21.1 t/s
2.79 s
—
猫羽霖API
L1
98%
DeepSeekdeepseek-ai/deepseek-v4-flash
2api
$0.140/M
Cache read$0.028/MCache write$1.12/MCache write 1h$1.79/M
$0.280/M
16.3 t/s
14.96 s
788280100
L1
98%
DeepSeekdeepseek-v4-flash-free
懒人
-98%$0.0025/M
Cache read$0.020/MCache write$0.020/MCache write 1h$0.032/M
-97%$0.0084/M
—
—
—
L1
98%
DeepSeekdeepseek/deepseek-v4-flash
懒人
$0.140/M
Cache read$0.028/MCache write$1.12/MCache write 1h$1.79/M
$0.280/M
—
—
—
Showing 20 model IDs of 178.

Alternatives & Similar Models

DeepSeekDeepSeek V4 Pro

deepseek-v4-pro

DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.

247 shared providers

OpenAIGPT-5.4

gpt-5-4

OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.

212 shared providers

ChatGLMGLM-5.1

glm-5-1

Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.

210 shared providers

MoonshotAIKimi K2.5

kimi-k2-5

Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.

197 shared providers

DeepSeekDeepSeek V3.2

deepseek-v3-2

DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.

191 shared providers

ClaudeClaude Sonnet 4.6

claude-sonnet-4-6

Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.

188 shared providers

Related model comparisons

DDeepSeek V4 Flash vs DeepSeek V4 Pro

6 verified comparison points

Indexable

DDeepSeek V4 Flash vs GPT-5.4

6 verified comparison points

Indexable

DDeepSeek V4 Flash vs GLM-5.1

6 verified comparison points

Indexable

Frequently Asked Questions

What benchmark data does DeepSeek V4 Flash include?
LMSpeed shows DeepSeek V4 Flash benchmark context, API price, output speed, first-token latency, and provider data across 381 providers when those signals are available.
What is the DeepSeek V4 Flash API price?
DeepSeek V4 Flash has pricing from 381 providers, ranging from $0.0000014/M to $1398.60/M. AIGCBAR has the lowest listed price.
What does the DeepSeek V4 Flash API pricing table include?
The DeepSeek V4 Flash API pricing table compares 381 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest DeepSeek V4 Flash API pricing?
AIGCBAR currently has the lowest listed DeepSeek V4 Flash price at $0.0000014/M across 381 providers.
Can I compare DeepSeek V4 Flash API price and speed together?
Yes. LMSpeed shows DeepSeek V4 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is DeepSeek V4 Flash API free?
Yes, DeepSeek V4 Flash free API options are available through 9 providers on LMSpeed, including 初叶🍂Furry API, 兔子API, Dext API, Moyanjdc API, 91VIP API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get DeepSeek V4 Flash free API access?
LMSpeed currently lists 9 free API providers for DeepSeek V4 Flash: 初叶🍂Furry API, 兔子API, Dext API, Moyanjdc API, 91VIP API. Check each provider row before using it because free tier limits can change.

Also known as

DeepSeek-V4-FlashDeepSeek-V4-Flash-2026-04-23DeepSeek-V4-Flash-fastDeepSeek-V4-Flash-globalDeepSeek-V4-flash

Data as of Aug 10, 2026, 07:34 PM·Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.·Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation