llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsDeepSeek-R1-Distill-Llama

DeepSeek model product

DeepSeek-R1-Distill-Llama

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token).

Updated Aug 12, 2026. Default version: DeepSeek R1 Distill Llama 8B

Compare
LLMBoard score12.4DeepSeek R1 Distill Llama 8B
Coverage80%4 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

DeepSeek-R1-Distill-Llama Specifications

Technical details for the model's default version.

Version
DeepSeek R1 Distill Llama 8B
Released
Jan 20, 2025
Knowledge cutoff
Unknown
Parameters
8B
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
No
License
MIT

DeepSeek-R1-Distill-Llama Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek R1 Distill Llama 8B category scores

DeepSeek-R1-Distill-Llama Benchmark Results

Benchmark scores for DeepSeek R1 Distill Llama 8B.

4 rows
Columns

Show columns

AIME 20240.8295346.1%CAug 11, 2026
MATH-5000.929329.7%CAug 11, 2026
LiveCodeBench0.4507331.9%CAug 11, 2026
GPQA0.518123422.8%CAug 11, 2026

DeepSeek-R1-Distill-Llama Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

DeepSeek-R1-Distill-Llama Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
1
1 rows
Columns

Show columns

Alibaba (China)deepseek-r1-distill-llama-8bglobalN/AN/A32.8KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-R1-Distill-Llama Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

DeepSeek R1 Distill Llama 70BJan 20, 2025N/A70.6B128K128KNoMIT
DeepSeek R1 Distill Llama 8BJan 20, 202512.48BN/AN/ANoMIT

What is DeepSeek-R1-Distill-Llama?

Key information about DeepSeek-R1-Distill-Llama and its available data.

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

Data as of 2026-08-11.

DeepSeek-R1-Distill-Llama vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-R1-Distill-LlamavsNova LiteDeepSeek-R1-Distill-LlamavsLlama 3.1 70BDeepSeek-R1-Distill-LlamavsQwen2.5 VL 7BDeepSeek-R1-Distill-LlamavsGemma 4 E2BDeepSeek-R1-Distill-LlamavsGemma 3 12BDeepSeek-R1-Distill-LlamavsLlama 3.2 90B

Models similar to DeepSeek-R1-Distill-Llama

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#206+2.8
DE

DeepSeek-V2.5

DeepSeek

15.2 LLMBoard

DetailsCompare
#205+3.3
DE

DeepSeek-R1-Distill-Qwen

DeepSeek

15.7 LLMBoard

DetailsCompare
#247-12.4
DE

DeepSeek-VL2

DeepSeek

0.0 LLMBoard

DetailsCompare
#169+13.6
DE

DeepSeek-V3

DeepSeek

26.0 LLMBoard

DetailsCompare
#128+27.5
DE

DeepSeek-V3.1

DeepSeek

39.9 LLMBoard

DetailsCompare
#111+31.9
DE

DeepSeek-R1

DeepSeek

44.2 LLMBoard

DetailsCompare

FAQ

Common questions about DeepSeek-R1-Distill-Llama.

When was DeepSeek-R1-Distill-Llama released?

DeepSeek-R1-Distill-Llama's default version was released on Jan 20, 2025.

How much does DeepSeek-R1-Distill-Llama cost?

No official standard PAYG price is currently available for DeepSeek-R1-Distill-Llama.

Who created DeepSeek-R1-Distill-Llama?

DeepSeek-R1-Distill-Llama was created by DeepSeek.

What is the context window for DeepSeek-R1-Distill-Llama?

A context window is not available for the default version.

Is DeepSeek-R1-Distill-Llama open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer DeepSeek-R1-Distill-Llama?

1 provider offerings are linked to the default version.

What models should I compare DeepSeek-R1-Distill-Llama with?

Nearby ranked alternatives include Nova Lite, Llama 3.1 70B, Qwen2.5 VL 7B.

Rankings

OverallCodingText ArenaPricing

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai