DeepSeek V3 0324 API Benchmarks, Pricing & Provider Data
Compare DeepSeek V3 0324 with another model
Choose a model to open its comparison page.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) mod...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Mar 2025
- Knowledge cutoff
- 2024-07-31
- Tokenizer
- DeepSeek
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
5 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Crusoe crusoe/bf16 | $0.500/M | $1.50/M | 100.0% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $0.290/M | $1.14/M | 99.1% | — | — | undefined tokens / — |
Novita novita/fp8 | $0.270/M | $1.12/M | 98.4% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.250/M | $1/M | 95.5% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp4 | $0.240/M | $0.900/M | 92.1% | — | — | undefined tokens / undefined tokens |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
