DeepSeek V4 Flash API Benchmarks, Pricing & Provider Data
Compare DeepSeek V4 Flash with another model
Choose a model to open its comparison page.
DeepSeek V4 Flash benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0000044/M. DeepSeek V4 Flash free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reason...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 5 / 8
- Methodology
- V3.0
#1Reasoning53.3RatedGlobal rank #303/4 Measured dimensions
#2Agents45.4RatedGlobal rank #514/4 Measured dimensions
#3Coding43.9RatedGlobal rank #374/4 Measured dimensions
#4Knowledge41.4Provisional1/4 Measured dimensions
#5Math32.6Estimated2/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Apr 2026
- Tokenizer
- DeepSeek
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_completion_tokensmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score51.0#79 / 112
Overall
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.130/M#18 / 186Output price$0.280/M#10 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score45.4#5180% interval39.1–51.74/4 Measured dimensionsAgentic score34.6#65 / 77Terminal-Bench 2.056.6#38 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score43.9#3780% interval37.3–50.54/4 Measured dimensionsSciCode50.3%#32 / 89Coding score46.0#65 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Score53.3#3080% interval44.3–62.33/4 Measured dimensionsMMLU-Pro86.4%#17 / 129GPQA86.7%#56 / 218
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score41.480% interval23.4–59.51/4 Measured dimensionsKnowledge score52.8#69 / 83SimpleQA28.9#4 / 4
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score32.680% interval20.4–44.82/4 Measured dimensionsHMMT Feb 202691.9#8 / 18IMOAnswerBench85.1#6 / 7
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1230.0#43 / 78
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
17 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Novita novita/fp8 | $0.140/M | $0.280/M | 100.0% | — | — | undefined tokens / undefined tokens |
Wafer wafer | $0.100/M | $0.250/M | 100.0% | — | — | undefined tokens / undefined tokens |
Phala phala | $0.200/M | $0.400/M | 100.0% | — | — | undefined tokens / undefined tokens |
Baidu baidu/fp8 | $0.066/M | $0.132/M | 100.0% | — | — | undefined tokens / undefined tokens |
AtlasCloud atlas-cloud/fp4 | $0.140/M | $0.280/M | 99.9% | — | — | undefined tokens / undefined tokens |
NextBit nextbit/fp8 | $0.150/M | $0.350/M | 99.9% | — | — | undefined tokens / undefined tokens |
DigitalOcean digitalocean | $0.068/M | $0.168/M | 99.9% | — | — | undefined tokens / undefined tokens |
Parasail parasail/fp8 | $0.140/M | $0.280/M | 99.9% | — | — | undefined tokens / undefined tokens |
Alibaba alibaba/fp8 | $0.134/M | $0.268/M | 99.8% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $0.091/M | $0.182/M | 99.8% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.130/M | $0.280/M | 99.6% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp8 | $0.090/M | $0.180/M | 99.4% | — | — | undefined tokens / undefined tokens |
Venice venice | $0.097/M | $0.193/M | 99.2% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake/fp8 | $0.066/M | $0.131/M | 98.8% | — | — | undefined tokens / undefined tokens |
Azure azure/us | $0.210/M | $0.560/M | 98.6% | — | — | undefined tokens / undefined tokens |
Mancer 2 mancer/fp8 | $0.190/M | $0.500/M | 97.6% | — | — | undefined tokens / undefined tokens |
OpenInference open-inference/fp8 | $0.050/M | $0.120/M | 93.9% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare DeepSeek V4 Flash API pricing across 386 providers. Prices range from $0.0000044/M to $1071.43/M. AIGCBAR offers the lowest rate at $0.0000044/M. 22 providers offer free API credits or a free tier.
Alternatives & Similar Models
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Frequently Asked Questions
- What benchmark data does DeepSeek V4 Flash include?
- LMSpeed shows DeepSeek V4 Flash benchmark context, API price, output speed, first-token latency, and provider data across 408 providers when those signals are available.
- What is the DeepSeek V4 Flash API price?
- DeepSeek V4 Flash has pricing from undefined provider} other undefined providers}}, ranging from $0.0000044/M to $1071.43/M. AIGCBAR has the lowest listed price.
- What does the DeepSeek V4 Flash API pricing table include?
- The DeepSeek V4 Flash API pricing table compares 408 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest DeepSeek V4 Flash API pricing?
- AIGCBAR currently has the lowest listed DeepSeek V4 Flash price at $0.0000044/M across undefined provider} other undefined providers}}.
- Can I compare DeepSeek V4 Flash API price and speed together?
- Yes. LMSpeed shows DeepSeek V4 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is DeepSeek V4 Flash API free?
- Yes, DeepSeek V4 Flash free API options are available through 22 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, DeadlySignal API, DeadlySignal API, DeadlySignal API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get DeepSeek V4 Flash free API access?
- LMSpeed currently lists 22 free API providerundefined other undefined} for DeepSeek V4 Flash: 兔子API, 兔子API, DeadlySignal API, DeadlySignal API, DeadlySignal API. Check each provider row before using it because free tier limits can change.
