llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsDeepSeek-R1-Distill-Llama

DeepSeek model product

DeepSeek-R1-Distill-Llama

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

Updated Aug 10, 2026. Default version: DeepSeek R1 Distill Llama 8B

Compare
LLMBoard scoreN/ANot scored
CoverageN/ANo published score
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
DeepSeek R1 Distill Llama 8B
Released
Jan 20, 2025
Knowledge cutoff
Unknown
Parameters
8B
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
No
License
MIT

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

No capability profile

No published version under this unique model currently has enough benchmark coverage to calculate a score.

Benchmark results

Published benchmark records for the scored version currently unavailable.

No benchmark rows

No published benchmark result is linked to the scored version.

Arena results

Preference and agent-evaluation signals from published Arena datasets.

No published Arena match

The default version has no published Arena rows, or its source alias has not been resolved.

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
1
1 rows
Columns

Show columns

Alibaba (China)deepseek-r1-distill-llama-8bglobalN/AN/A32.8KAug 7, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-R1-Distill-Llama versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

2 rows
Columns

Show columns

DeepSeek R1 Distill Llama 70BJan 20, 2025N/A70.6B128K128KNoMIT
DeepSeek R1 Distill Llama 8BJan 20, 2025N/A8BN/AN/ANoMIT

What is DeepSeek-R1-Distill-Llama?

A concise description based on the published model registry.

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Models similar to DeepSeek-R1-Distill-Llama

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#126
DE

DeepSeek-V3.1

DeepSeek

36.0 LLMBoard

DetailsCompare
#119
DE

DeepSeek-V3

DeepSeek

38.1 LLMBoard

DetailsCompare
#101
DE

DeepSeek-R1

DeepSeek

44.1 LLMBoard

DetailsCompare
#32
DE

DeepSeek-V2.5

DeepSeek

69.5 LLMBoard

DetailsCompare
#28
DE

DeepSeek-V4-Flash

DeepSeek

71.7 LLMBoard

DetailsCompare
#18
DE

DeepSeek-V3.2

DeepSeek

75.5 LLMBoard

DetailsCompare

FAQ

Common questions about DeepSeek-R1-Distill-Llama.

When was DeepSeek-R1-Distill-Llama released?

DeepSeek-R1-Distill-Llama's default version was released on Jan 20, 2025.

How much does DeepSeek-R1-Distill-Llama cost?

No official standard PAYG price is currently available for DeepSeek-R1-Distill-Llama.

Who created DeepSeek-R1-Distill-Llama?

DeepSeek-R1-Distill-Llama is published under DeepSeek in the model registry.

What is the context window for DeepSeek-R1-Distill-Llama?

The current registry does not publish a context window for the default version.

Is DeepSeek-R1-Distill-Llama open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer DeepSeek-R1-Distill-Llama?

1 published provider offerings are linked to the default version.

What models should I compare DeepSeek-R1-Distill-Llama with?

No nearby ranked alternatives are currently available.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai