#126
DE
DeepSeek-V3.1
DeepSeek
36.0 LLMBoard
DeepSeek model product
DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.
Updated Aug 10, 2026. Default version: DeepSeek R1 Distill Qwen 7B
Structured fields from the published default version.
This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.
No published version under this unique model currently has enough benchmark coverage to calculate a score.
Published benchmark records for the scored version currently unavailable.
No published benchmark result is linked to the scored version.
Preference and agent-evaluation signals from published Arena datasets.
The default version has no published Arena rows, or its source alias has not been resolved.
Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.
| Alibaba (China) | deepseek-r1-distill-qwen-7b | global | $0.072 | $0.144 | 32.8K |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.
| DeepSeek R1 Distill Qwen 1.5B | N/A | 1.8B | N/A | N/A | No | MIT | |
| DeepSeek R1 Distill Qwen 14B | N/A | 14.8B | N/A | N/A | No | MIT | |
| DeepSeek R1 Distill Qwen 32B | N/A | 32.8B | 128K | 128K | No | MIT | |
| DeepSeek R1 Distill Qwen 7B | N/A | 7.6B | N/A | N/A | No | MIT |
A concise description based on the published model registry.
DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.
Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.
Data snapshot: 2026-08-07. Editorial model content is not available in the backend.
Recommendations prioritize the same model type and family, then the closest published LLMBoard score.
Common questions about DeepSeek-R1-Distill-Qwen.
DeepSeek-R1-Distill-Qwen's default version was released on Jan 20, 2025.
No official standard PAYG price is currently available for DeepSeek-R1-Distill-Qwen. The lowest tracked third-party offer starts at $0.072 input and $0.144 output via Alibaba (China).
DeepSeek-R1-Distill-Qwen is published under DeepSeek in the model registry.
The current registry does not publish a context window for the default version.
No. The default version is not marked as having publicly available weights.
1 published provider offerings are linked to the default version.
No nearby ranked alternatives are currently available.