llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

DeepSeek model product

DeepSeek-R1-Distill-Llama

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token).

Updated Aug 12, 2026. Default version: DeepSeek R1 Distill Llama 8B

Compare
LLMBoard score12.4DeepSeek R1 Distill Llama 8B
Coverage60%4 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

DeepSeek-R1-Distill-Llama Specifications

Technical details for the model's default version.

Version
DeepSeek R1 Distill Llama 8B
Released
Jan 20, 2025
Knowledge cutoff
Unknown
Parameters
8B
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
No
License
MIT

DeepSeek-R1-Distill-Llama Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek R1 Distill Llama 8B category scores

DeepSeek-R1-Distill-Llama Benchmark Results

Benchmark scores for DeepSeek R1 Distill Llama 8B.

4 rows
Columns

Show columns

AIME 202480.0%295346.1%CAug 11, 2026
MATH-50089.1%29329.7%CAug 11, 2026
LiveCodeBench39.6%507331.9%CAug 11, 2026
GPQA49.0%18123422.8%CAug 11, 2026

DeepSeek-R1-Distill-Llama Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

DeepSeek-R1-Distill-Llama Runtime Performance

Provider-specific output speed and catalog latency for DeepSeek R1 Distill Llama 8B. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

DeepSeek-R1-Distill-Llama Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
1
1 rows
Columns

Show columns

Alibaba (China)deepseek-r1-distill-llama-8bglobalN/AN/A32.8KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-R1-Distill-Llama Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

DeepSeek R1 Distill Llama 70BJan 20, 2025N/A70.6B128K128KNoMIT
DeepSeek R1 Distill Llama 8BJan 20, 202512.48BN/AN/ANoMIT

What is DeepSeek-R1-Distill-Llama?

Key information about DeepSeek-R1-Distill-Llama and its available data.

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

Data as of 2026-08-11.

DeepSeek-R1-Distill-Llama vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-R1-Distill-LlamavsNova LiteDeepSeek-R1-Distill-LlamavsLlama 3.1 70BDeepSeek-R1-Distill-LlamavsQwen2.5 VL 7BDeepSeek-R1-Distill-LlamavsGemma 4 E2BDeepSeek-R1-Distill-LlamavsGemma 3 12BDeepSeek-R1-Distill-LlamavsLlama 3.2 90B

Models similar to DeepSeek-R1-Distill-Llama

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#206+3.1
DE

DeepSeek-V2.5

DeepSeek

15.4 LLMBoard

DetailsCompare
#205+3.3
DE

DeepSeek-R1-Distill-Qwen

DeepSeek

15.7 LLMBoard

DetailsCompare
#247-12.4
DE

DeepSeek-VL2

DeepSeek

0.0 LLMBoard

DetailsCompare
#169+13.6
DE

DeepSeek-V3

DeepSeek

26.0 LLMBoard

DetailsCompare
#128+27.5
DE

DeepSeek-V3.1

DeepSeek

39.9 LLMBoard

DetailsCompare
#111+31.9
DE

DeepSeek-R1

DeepSeek

44.2 LLMBoard

DetailsCompare

FAQ

Common questions about DeepSeek-R1-Distill-Llama.

When was DeepSeek-R1-Distill-Llama released?

DeepSeek-R1-Distill-Llama's default version was released on Jan 20, 2025.

How much does DeepSeek-R1-Distill-Llama cost?

No official standard PAYG price is currently available for DeepSeek-R1-Distill-Llama.

Who created DeepSeek-R1-Distill-Llama?

DeepSeek-R1-Distill-Llama was created by DeepSeek.

What is the context window for DeepSeek-R1-Distill-Llama?

A context window is not available for the default version.

Is DeepSeek-R1-Distill-Llama open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer DeepSeek-R1-Distill-Llama?

1 provider offerings are linked to the default version.

What models should I compare DeepSeek-R1-Distill-Llama with?

Nearby ranked alternatives include Nova Lite, Llama 3.1 70B, Qwen2.5 VL 7B.