llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

DeepSeek model product

DeepSeek-R1-Distill-Llama

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token).

Updated Aug 17, 2026. Default version: DeepSeek R1 Distill Llama 8B

Compare
LLMBoard Score11.6DeepSeek R1 Distill Llama 8B
Coverage60%4 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

DeepSeek-R1-Distill-Llama Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek R1 Distill Llama 8B LLMBoard score breakdown

DeepSeek-R1-Distill-Llama Benchmark Results

Benchmark scores for DeepSeek R1 Distill Llama 8B.

4 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkAIME 2024Score80.0%Rank29Participants53Percentile46.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkMATH-500Score89.1%Rank29Participants32Percentile9.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBenchScore39.6%Rank52Participants75Percentile31.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore49.0%Rank186Participants239Percentile22.3%EvidenceCEvaluatedAug 17, 2026

DeepSeek-R1-Distill-Llama Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

DeepSeek-R1-Distill-Llama Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
1
1 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba (China)Provider model IDdeepseek-r1-distill-llama-8bRegionglobalInput / 1MN/AOutput / 1MN/AContext32.8KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-R1-Distill-Llama Runtime Performance

Provider-specific output speed and catalog latency for DeepSeek R1 Distill Llama 8B. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

DeepSeek-R1-Distill-Llama Specifications

Technical details for the model's default version.

Version
DeepSeek R1 Distill Llama 8B
Released
Jan 20, 2025
Knowledge cutoff
Unknown
Parameters
8B
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
No
License
MIT

DeepSeek-R1-Distill-Llama Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionDeepSeek R1 Distill Llama 70BReleasedJan 20, 2025LLMBoardN/AParameters70.6BContext128KMax output128KOpen weightsNoLicenseMIT
VersionDeepSeek R1 Distill Llama 8BReleasedJan 20, 2025LLMBoard11.6Parameters8BContextN/AMax outputN/AOpen weightsNoLicenseMIT

DeepSeek-R1-Distill-Llama vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-R1-Distill-LlamavsNova LiteDeepSeek-R1-Distill-LlamavsLlama 3.1 70BDeepSeek-R1-Distill-LlamavsQwen2.5 VL 7BDeepSeek-R1-Distill-LlamavsGemma 4 E2BDeepSeek-R1-Distill-LlamavsGemini 1.5 FlashDeepSeek-R1-Distill-LlamavsGemma 3 12B

Models similar to DeepSeek-R1-Distill-Llama

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#221+2.2
DE

DeepSeek-V2.5

DeepSeek

13.8 LLMBoard

DetailsCompare
#217+3.3
DE

DeepSeek-R1-Distill-Qwen

DeepSeek

14.8 LLMBoard

DetailsCompare
#258-11.6
DE

DeepSeek-VL2

DeepSeek

0.0 LLMBoard

DetailsCompare
#181+14.0
DE

DeepSeek-V3

DeepSeek

25.5 LLMBoard

DetailsCompare
#138+27.6
DE

DeepSeek-V3.1

DeepSeek

39.2 LLMBoard

DetailsCompare
#122+31.9
DE

DeepSeek-R1

DeepSeek

43.5 LLMBoard

DetailsCompare

What is DeepSeek-R1-Distill-Llama?

Key information about DeepSeek-R1-Distill-Llama and its available data.

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

Data as of 2026-08-17.

FAQ

Common questions about DeepSeek-R1-Distill-Llama.

When was DeepSeek-R1-Distill-Llama released?

DeepSeek-R1-Distill-Llama's default version was released on Jan 20, 2025.

How much does DeepSeek-R1-Distill-Llama cost?

No official standard PAYG price is currently available for DeepSeek-R1-Distill-Llama.

Who created DeepSeek-R1-Distill-Llama?

DeepSeek-R1-Distill-Llama was created by DeepSeek.

What is the context window for DeepSeek-R1-Distill-Llama?

A context window is not available for the default version.

Is DeepSeek-R1-Distill-Llama open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer DeepSeek-R1-Distill-Llama?

1 provider offerings are linked to the default version.

What models should I compare DeepSeek-R1-Distill-Llama with?

Nearby ranked alternatives include Nova Lite, Llama 3.1 70B, Qwen2.5 VL 7B.