DeepSeek model product
A powerful Mixture-of-Experts (MoE) language model with 671B total parameters (37B activated per token).
Updated Aug 12, 2026. Default version: DeepSeek-V3 0324
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for DeepSeek-V3 0324.
| MATH-500 | 0.9 | 22 | 32 | 32.3% | C | |
| AIME 2024 | 0.6 | 44 | 53 | 17.3% | C | |
| MMLU-Pro | 0.8 | 44 | 129 | 66.4% | C | |
| LiveCodeBench | 0.5 | 47 | 73 | 36.1% | C | |
| GPQA | 0.7 | 136 | 234 | 42.1% | C |
Preference and agent-evaluation results for the default version.
| text style control | overall | 142 | 1395.7 | 45,626 | N/A | |
| text | overall | 152 | 1375.1 | 45,626 | N/A |
Official vendor API pricing appears first, followed by individual provider offers.
| NanoGPT | deepseek-v3-0324 | global | $0.20 | $0.77 | 128K | |
| submodel | deepseek-ai/DeepSeek-V3-0324 | global | $0.20 | $0.80 | 75K | |
| Meganova | deepseek-ai/DeepSeek-V3-0324 | global | $0.25 | $0.88 | 163.8K | |
| NovitaAI | deepseek/deepseek-v3-0324 | global | $0.27 | $1.12 | 163.8K | |
| OpenRouter | deepseek/deepseek-chat-v3-0324 | global | $0.27 | $1.12 | 163.8K | |
| Jiekou.AI | deepseek/deepseek-v3-0324 | global | $0.28 | $1.14 | 163.8K |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| DeepSeek-V3 0324 | 26.0 | 671B | 163.8K | 163.8K | No | MIT + Model License (Commercial use allowed) | |
| DeepSeek-V3 | N/A | 671B | 131.1K | 8.2K | Yes | MIT + Model License (Commercial use allowed) |
Key information about DeepSeek-V3 and its available data.
A powerful Mixture-of-Experts (MoE) language model with 671B total parameters (37B activated per token). Features Multi-head Latent Attention (MLA), auxiliary-loss-free load balancing, and multi-token prediction training. 8T tokens with strong performance in reasoning, math, and code tasks.
Data as of 2026-08-11.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about DeepSeek-V3.
DeepSeek-V3's default version was released on Mar 25, 2025.
No official standard PAYG price is currently available for DeepSeek-V3. The lowest tracked third-party offer starts at $0.20 input and $0.77 output via NanoGPT.
DeepSeek-V3 was created by DeepSeek.
The default version has a 163.8K token context window.
No. The default version is not marked as having publicly available weights.
7 provider offerings are linked to the default version.
Nearby ranked alternatives include Qwen3 VL 8B Thinking, Gemini 2.0 Flash, Phi 4 Reasoning Plus.