DeepSeek
·Released on Jul 31, 2026

DeepSeek V4 Flash 0731 API Benchmarks, Pricing & Provider Data

Compare DeepSeek V4 Flash 0731 with another model

Choose a model to open its comparison page.

Share on X
LLM

DeepSeek V4 Flash 0731 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.000044/M. DeepSeek V4 Flash 0731 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workfl...

Quality
#55of 112
59.0
LMSpeed score
Speed
74char/s
2.70 s
Cost
$0.000044/ 1M · 8:1 in:out
$0.00000704 in · $0.000037 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
5 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents51.3Coding50.6Reasoning58.7Knowledge47.5Math57.9Multilingual-Multimodal-Instruction following-
#1Reasoning58.7RatedGlobal rank #93/4 Measured dimensions
80% interval: 50.367.0
abstract logicPrior only
scientific causal60.8
critpt · critpt / gpqa · aa_gpqa_diamond / hle · aa_hle
multistep constraints57.5
mmlu_pro · mmlu_pro
evidence verification57.7
lcr · lcr / mrcr · mrcr1m
#2Math57.9Estimated2/4 Measured dimensions
80% interval: 45.870.1
foundational mathPrior only
competition math60.7
hmmt · hmmt_feb2026
proof frontier55.2
imo · imo_answer_bench
applied tool mathPrior only
#3Agents51.3RatedGlobal rank #353/4 Measured dimensions
80% interval: 42.959.6
planningPrior only
tool use51
mcp_atlas · mcp_atlas / toolathlon · toolathlon
environment execution47.9
browsecomp · browse_comp / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability55.1
cyber_gym · cyber_gym
#4Coding50.6RatedGlobal rank #254/4 Measured dimensions
80% interval: 44.356.9
code generation56.2
scicode · aa_sci_code
repository engineering50.2
nl2repo · nl2_repo / swe_pro · swe_pro
debugging testing50.7
swe_multilingual · benchlm_coding_swe_multilingual / swe_verified · swe_verified
tooling quality45.1
terminalbench · benchlm_coding_terminal_bench2
#5Knowledge47.5Provisional1/4 Measured dimensions
80% interval: 29.565.6
broad knowledgePrior only
professional knowledgePrior only
factuality47.5
simpleqa · simple_qa
retrieval open bookPrior only
No data:MultilingualMultimodalInstruction following

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1.3Mtokens
1.6K pages of text
OUTPUT
943.7Ktokens
8K128K1M4M
1.3M

Features

Technical Details

Input
Output
Released
Jul 2026
Tokenizer
DeepSeek
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_pparallel_tool_callspresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_logprobstop_p

Rankings

Excels at

Falls behind in

Detailed scores

Updated: Aug 20, 2026

Overall

undefined metric} other undefined metrics}}

Overall score59.0#55 / 112DeepSWE54.4#16 / 21
Overall score
LMSpeed rank#55
Score59.0
UpdatedAug 20, 2026
Confidence2
DeepSWE
LMSpeed rank#16
Score54.4
UpdatedAug 20, 2026
Confidence2

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score51.3#3580% interval42.9–59.63/4 Measured dimensionsAgentic score49.3#57 / 77Terminal-Bench 2.056.9#37 / 53
Dimensions and evidence
Planning & decompositionPrior only
Tool use51
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z -0.40 · q 1.00
toolathlon · toolathlon · z -0.02 · q 1.00
Environment & long-horizon execution47.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -0.68 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z 0.06 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.43 · q 1.00
Recovery & completion reliability55.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
cyber_gym · cyber_gym · z -0.06 · q 1.00
Agentic score
LMSpeed rank#57
Score49.3
UpdatedAug 20, 2026
Confidence2
Terminal-Bench 2.0V3 evidence
LMSpeed rank#37
Score56.9
UpdatedAug 20, 2026
Confidence2
BrowseCompV3 evidence
LMSpeed rank#22
Score73.2
UpdatedAug 20, 2026
Confidence2
HLE w/ tools
LMSpeed rank#9
Score45.1
UpdatedAug 20, 2026
Confidence2
MCP AtlasV3 evidence
LMSpeed rank#20
Score69.0
UpdatedAug 20, 2026
Confidence2
GDPval-AAV3 evidence
LMSpeed rank#34
Score1189.0
UpdatedAug 20, 2026
Confidence2
ToolathlonV3 evidence
LMSpeed rank#14
Score47.8
UpdatedAug 20, 2026
Confidence2
AA Agentic Index
LMSpeed rank#11
Score48.4
UpdatedAug 20, 2026
Confidence2
GDPval-AA
LMSpeed rank#11
Score52.9
UpdatedAug 20, 2026
Confidence2
Terminal-Bench 2.1
LMSpeed rank#8
Score82.7
UpdatedAug 20, 2026
Confidence2
CyberGymV3 evidence
LMSpeed rank#8
Score76.7
UpdatedAug 20, 2026
Confidence2
Toolathlon-Verified
LMSpeed rank#7
Score70.3
UpdatedAug 20, 2026
Confidence2
Agents' Last Exam
LMSpeed rank#7
Score25.2
UpdatedAug 20, 2026
Confidence2
AutomationBench
LMSpeed rank#10
Score25.1
UpdatedAug 20, 2026
Confidence2

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score50.6#2580% interval44.3–56.94/4 Measured dimensionsCoding score53.4#43 / 87Codeforces3052.0#2 / 4
Dimensions and evidence
Code generation56.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · aa_sci_code · z 0.65 · q 1.00
Repository engineering50.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
nl2repo · nl2_repo · z 1.18 · q 1.00
swe_pro · swe_pro · z -0.78 · q 1.00
Debugging & testing50.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_multilingual · benchlm_coding_swe_multilingual · z -0.58 · q 1.00
swe_verified · swe_verified · z 0.45 · q 1.00
Tool-assisted development & quality45.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · benchlm_coding_terminal_bench2 · z -0.63 · q 1.00
Coding score
LMSpeed rank#43
Score53.4
UpdatedAug 20, 2026
Confidence2
Codeforces
LMSpeed rank#2
Score3052.0
UpdatedAug 20, 2026
Confidence2
SWE-bench VerifiedV3 evidence
LMSpeed rank#17
Score79.0
UpdatedAug 20, 2026
Confidence2
SWE-bench ProV3 evidence
LMSpeed rank#42
Score52.6
UpdatedAug 20, 2026
Confidence2
SWE MultilingualV3 evidence
LMSpeed rank#15
Score73.3
UpdatedAug 20, 2026
Confidence2
Terminal-Bench 2.0V3 evidence
LMSpeed rank#27
Score56.9
UpdatedAug 20, 2026
Confidence2
AA Coding Index
LMSpeed rank#24
Score69.1
UpdatedAug 20, 2026
Confidence2
AA-SciCodeV3 evidence
LMSpeed rank#36
Score49.9
UpdatedAug 20, 2026
Confidence2
LiveCodeBench Pass@1-COT
LMSpeed rank#2
Score91.6
UpdatedAug 20, 2026
Confidence2
Terminal-Bench 2.1
LMSpeed rank#8
Score82.7
UpdatedAug 20, 2026
Confidence2
NL2RepoV3 evidence
LMSpeed rank#4
Score54.2
UpdatedAug 20, 2026
Confidence2
DSBench-FullStack
LMSpeed rank#2
Score68.7
UpdatedAug 20, 2026
Confidence2
DSBench-Hard
LMSpeed rank#2
Score59.6
UpdatedAug 20, 2026
Confidence2

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score58.7#980% interval50.3–67.03/4 Measured dimensionsMMLU-Pro86.2%#18 / 129GPQA88.1%#49 / 218
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning60.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.70 · q 1.00
gpqa · aa_gpqa_diamond · z 0.97 · q 1.00
hle · aa_hle · z 0.83 · q 1.00
Multi-step constraints57.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 0.80 · q 1.00
Evidence integration & verification57.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.05 · q 1.00
mrcr · mrcr1m · z 1.23 · q 0.75
MMLU-ProV3 evidence
LMSpeed rank#18
Score86.2%
UpdatedAug 20, 2026
Confidence2
GPQAV3 evidence
LMSpeed rank#49
Score88.1%
UpdatedAug 20, 2026
Confidence2
HLEV3 evidence
LMSpeed rank#44
Score34.8%
UpdatedAug 20, 2026
Confidence2
MRCR 1MV3 evidence
LMSpeed rank#3
Score78.7
UpdatedAug 20, 2026
Confidence2
CorpusQA 1M
LMSpeed rank#2
Score60.5
UpdatedAug 20, 2026
Confidence2
AA-LCRV3 evidence
LMSpeed rank#52
Score74.3
UpdatedAug 20, 2026
Confidence2
CritPtV3 evidence
LMSpeed rank#24
Score16.6
UpdatedAug 20, 2026
Confidence2

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score47.580% interval29.5–65.61/4 Measured dimensionsKnowledge score61.1#57 / 83SimpleQA34.1#3 / 4
Dimensions and evidence
Broad knowledgePrior only
Professional knowledgePrior only
Factuality47.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
simpleqa · simple_qa · z -0.43 · q 0.17
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#57
Score61.1
UpdatedAug 20, 2026
Confidence2
SimpleQAV3 evidence
LMSpeed rank#3
Score34.1
UpdatedAug 20, 2026
Confidence2
Chinese-SimpleQA
LMSpeed rank#2
Score78.9
UpdatedAug 20, 2026
Confidence2
GPQA-D
LMSpeed rank#23
Score88.1
UpdatedAug 20, 2026
Confidence2
Artificial Analysis Intelligence Index
LMSpeed rank#8
Score51.8
UpdatedAug 20, 2026
Confidence2
AA-GPQA Diamond
LMSpeed rank#28
Score90.8
UpdatedAug 20, 2026
Confidence2
AA-HLE
LMSpeed rank#32
Score38.6
UpdatedAug 20, 2026
Confidence2
AA-Omniscience Accuracy
LMSpeed rank#33
Score40.4
UpdatedAug 20, 2026
Confidence2
AA-Omniscience Hallucination Rate
LMSpeed rank#13
Score91.7
UpdatedAug 20, 2026
Confidence2

Math

V3.0

undefined metric} other undefined metrics}} · Estimated

Score57.980% interval45.8–70.12/4 Measured dimensionsHMMT Feb 202694.8#3 / 18IMOAnswerBench88.4#3 / 7
Dimensions and evidence
Foundational mathPrior only
Competition math60.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
hmmt · hmmt_feb2026 · z 1.55 · q 1.00
Advanced proofs55.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
imo · imo_answer_bench · z 0.20 · q 0.88
Applied & tool-assisted mathPrior only
HMMT Feb 2026V3 evidence
LMSpeed rank#3
Score94.8
UpdatedAug 20, 2026
Confidence2
IMOAnswerBenchV3 evidence
LMSpeed rank#3
Score88.4
UpdatedAug 20, 2026
Confidence2
Apex
LMSpeed rank#3
Score33.0
UpdatedAug 20, 2026
Confidence2
Apex Shortlist
LMSpeed rank#2
Score85.7
UpdatedAug 20, 2026
Confidence2
Math score
LMSpeed rank#6
Score81.4
UpdatedAug 20, 2026
Confidence2

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · No data

80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1219.0#45 / 78
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
Design Arena Website
LMSpeed rank#45
Score1219.0
UpdatedAug 20, 2026
Confidence2

Instruction following

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalizationPrior only
Multi-turn & long instructionsPrior only

OpenRouter endpoints

26 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Novita
novita/fp8
$0.409/M$1.23/M100.0%undefined tokens / undefined tokens
Alibaba
alibaba
$0.352/M$1.06/M100.0%undefined tokens / undefined tokens
Cloudflare
cloudflare
$0.440/M$1.32/M100.0%undefined tokens / undefined tokens
CoreWeave
coreweave/fp8
$0.130/M$0.280/M100.0%undefined tokens / undefined tokens
BaseTen
baseten/fp8
$0.130/M$0.260/M100.0%undefined tokens / undefined tokens
Morph
morph/bf16
$0.123/M$0.348/M99.9%undefined tokens / undefined tokens
AtlasCloud
atlas-cloud/fp4
$0.440/M$1.32/M99.9%undefined tokens / undefined tokens
GMICloud
gmicloud/fp8
$0.286/M$0.858/M99.9%undefined tokens / undefined tokens
DigitalOcean
digitalocean
$0.119/M$0.238/M99.8%undefined tokens / undefined tokens
Phala
phala
$0.440/M$1.32/M99.8%undefined tokens / undefined tokens
NextBit
nextbit/fp8
$0.352/M$1.06/M99.8%undefined tokens / undefined tokens
Inceptron
inceptron/fp4
$0.064/M$0.173/M99.8%undefined tokens / undefined tokens
Relace
relace/fp4
$0.060/M$0.120/M99.7%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.060/M$0.180/M99.7%undefined tokens / undefined tokens
Reka
reka/fp4
$0.110/M$0.660/M99.6%undefined tokens / undefined tokens
Baidu
baidu/fp8
$0.440/M$1.32/M99.6%undefined tokens / undefined tokens
Venice
venice
$0.175/M$0.350/M99.4%undefined tokens / undefined tokens
Wafer
wafer/fast
$0.100/M$0.250/M99.1%undefined tokens / undefined tokens
Sail Research
sail-research/fp4
$0.074/M$0.342/M99.0%undefined tokens / undefined tokens
OpenInference
open-inference/fp8
$0.040/M$0.100/M98.9%undefined tokens / undefined tokens
Fireworks
fireworks
$0.220/M$0.660/M98.2%undefined tokens / undefined tokens
Makora
makora
$0.090/M$0.195/M98.2%undefined tokens / undefined tokens
Mancer 2
mancer/fp8
$0.200/M$0.600/M98.1%undefined tokens / undefined tokens
Together
together
$0.140/M$0.280/M97.8%undefined tokens / undefined tokens
StreamLake
streamlake/fp8
$0.057/M$0.172/M97.1%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.220/M$0.660/M91.6%undefined tokens / undefined tokens

Pricing Comparison

Compare DeepSeek V4 Flash 0731 API pricing across 97 providers. Prices range from $0.000044/M to $10273.97/M. OAI2API offers the lowest rate at $0.000044/M. 4 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
L2
0%
deepseek/deepseek-v4-flash-0731
default
$0.112/M
Cache read$0.022/M
$0.224/M
74.1 t/s
2.70 s
L1
100%
deepseek-v4-flash-0731
deepseek
$0.500/M
$1.00/M
L1
100%
deepseek-v4-flash-0731
官转
$0.616/M
Cache read$0.021/M
$1.85/M
L1
100%
deepseek-v4-flash-0731
default
$0.411/M
Cache read$0.014/M
$1.23/M
L1
100%
deepseek-v4-flash-0731
default
$0.137/M
Cache read$0.014/M
$0.274/M
L1
100%
deepseek-v4-flash-0731
default
$0.205/M
Cache read$0.0068/M
$0.616/M
L1
100%
deepseek-ai/deepseek-v4-flash-0731
level1
$0.0010/request
-
L1
100%
deepseek-v4-flash-0731
Self-Deployed-2
$0.065/M
Cache read$0.0021/M
$0.196/M
L1
100%
deepseek-v4-flash-0731
opencode
$1.20/M
Cache read$0.040/M
$3.60/M
L1
100%
deepseek-v4-flash-0731
deepseek
$4.11/M
$8.22/M
L1
100%
deepseek-v4-flash-0731
default
$0.103/M
$0.925/M
L1
100%
L2
100%
deepseek-v4-flash-0731
diamond-glm
$0.292/M
Cache read$0.029/M
$0.949/M
L1
100%
deepseek-ai/DeepSeek-V4-Flash-0731
default
$7.50/M
$7.50/M
L1
100%
deepseek-ai/deepseek-v4-flash-0731
default
$7.50/M
$7.50/M
L1
100%
[量]deepseek-v4-flash-0731
default
$75.00/M
$75.00/M
L1
100%
deepseek-ai/deepseek-v4-flash-0731
NVIDIA英伟达
$0.00005/request
-
L1
100%
DeepSeek-V4-Flash-0731
国产模型
$0.700/M
Cache read$0.100/M
$1.00/M
L1
100%
L2
73%
deepseek-v4-flash-0731
default
$1.50/M
Cache read$0.050/M
$4.50/M
L1
100%
按次-deepseek-v4-flash-0731
按次计费
$0.065/request
-
L1
100%
deepseek-v4-flash-0731
Self-Deployed-2
$0.065/M
Cache read$0.0021/M
$0.196/M
Showing 20 model IDs of 55.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does DeepSeek V4 Flash 0731 include?
LMSpeed shows DeepSeek V4 Flash 0731 benchmark context, API price, output speed, first-token latency, and provider data across 101 providers when those signals are available.
What is the DeepSeek V4 Flash 0731 API price?
DeepSeek V4 Flash 0731 has pricing from undefined provider} other undefined providers}}, ranging from $0.000044/M to $10273.97/M. OAI2API has the lowest listed price.
What does the DeepSeek V4 Flash 0731 API pricing table include?
The DeepSeek V4 Flash 0731 API pricing table compares 101 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest DeepSeek V4 Flash 0731 API pricing?
OAI2API currently has the lowest listed DeepSeek V4 Flash 0731 price at $0.000044/M across undefined provider} other undefined providers}}.
Can I compare DeepSeek V4 Flash 0731 API price and speed together?
Yes. LMSpeed shows DeepSeek V4 Flash 0731 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is DeepSeek V4 Flash 0731 API free?
Yes, DeepSeek V4 Flash 0731 free API options are available through 4 providerundefined other undefined} on LMSpeed, including Dext API, 初叶🍂Furry API, 初叶🍂Furry API, 初叶🍂Furry API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get DeepSeek V4 Flash 0731 free API access?
LMSpeed currently lists 4 free API providerundefined other undefined} for DeepSeek V4 Flash 0731: Dext API, 初叶🍂Furry API, 初叶🍂Furry API, 初叶🍂Furry API. Check each provider row before using it because free tier limits can change.

Also known as

DeepSeek-V4-Flash-0731[量]deepseek-v4-flash-0731accounts/fireworks/models/deepseek-v4-flash-0731deepseek-ai/DeepSeek-V4-Flash-0731deepseek-ai/deepseek-v4-flash-0731

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation