llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

DeepSeek model product

DeepSeek-R1-Distill-Qwen

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token).

Updated Aug 12, 2026. Default version: DeepSeek R1 Distill Qwen 7B

Compare
LLMBoard score15.7DeepSeek R1 Distill Qwen 7B
Coverage60%4 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

DeepSeek-R1-Distill-Qwen Specifications

Technical details for the model's default version.

Version
DeepSeek R1 Distill Qwen 7B
Released
Jan 20, 2025
Knowledge cutoff
Unknown
Parameters
7.6B
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
No
License
MIT

DeepSeek-R1-Distill-Qwen Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek R1 Distill Qwen 7B category scores

DeepSeek-R1-Distill-Qwen Benchmark Results

Benchmark scores for DeepSeek R1 Distill Qwen 7B.

4 rows
Columns

Show columns

AIME 202483.3%215361.5%CAug 11, 2026
MATH-50092.8%243225.8%CAug 11, 2026
LiveCodeBench37.6%517330.6%CAug 11, 2026
GPQA49.1%18023423.2%CAug 11, 2026

DeepSeek-R1-Distill-Qwen Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

DeepSeek-R1-Distill-Qwen Runtime Performance

Provider-specific output speed and catalog latency for DeepSeek R1 Distill Qwen 7B. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

DeepSeek-R1-Distill-Qwen Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.072 input, $0.144 output per 1M via Alibaba (China)
Tracked offerings
1
1 rows
Columns

Show columns

Alibaba (China)deepseek-r1-distill-qwen-7bglobal$0.072$0.14432.8KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-R1-Distill-Qwen Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

4 rows
Columns

Show columns

DeepSeek R1 Distill Qwen 1.5BJan 20, 2025N/A1.8BN/AN/ANoMIT
DeepSeek R1 Distill Qwen 14BJan 20, 2025N/A14.8BN/AN/ANoMIT
DeepSeek R1 Distill Qwen 32BJan 20, 2025N/A32.8B128K128KNoMIT
DeepSeek R1 Distill Qwen 7BJan 20, 202515.77.6BN/AN/ANoMIT

What is DeepSeek-R1-Distill-Qwen?

Key information about DeepSeek-R1-Distill-Qwen and its available data.

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

Data as of 2026-08-11.

DeepSeek-R1-Distill-Qwen vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-R1-Distill-QwenvsMistral Large 3DeepSeek-R1-Distill-QwenvsGemma 3 27BDeepSeek-R1-Distill-QwenvsNova 2 SonicDeepSeek-R1-Distill-QwenvsDeepSeek-V2.5DeepSeek-R1-Distill-QwenvsQvQ 72BDeepSeek-R1-Distill-QwenvsQwen2.5 32B

Models similar to DeepSeek-R1-Distill-Qwen

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#206-0.2
DE

DeepSeek-V2.5

DeepSeek

15.4 LLMBoard

DetailsCompare
#215-3.3
DE

DeepSeek-R1-Distill-Llama

DeepSeek

12.4 LLMBoard

DetailsCompare
#169+10.3
DE

DeepSeek-V3

DeepSeek

26.0 LLMBoard

DetailsCompare
#247-15.7
DE

DeepSeek-VL2

DeepSeek

0.0 LLMBoard

DetailsCompare
#128+24.2
DE

DeepSeek-V3.1

DeepSeek

39.9 LLMBoard

DetailsCompare
#111+28.6
DE

DeepSeek-R1

DeepSeek

44.2 LLMBoard

DetailsCompare

FAQ

Common questions about DeepSeek-R1-Distill-Qwen.

When was DeepSeek-R1-Distill-Qwen released?

DeepSeek-R1-Distill-Qwen's default version was released on Jan 20, 2025.

How much does DeepSeek-R1-Distill-Qwen cost?

No official standard PAYG price is currently available for DeepSeek-R1-Distill-Qwen. The lowest tracked third-party offer starts at $0.072 input and $0.144 output via Alibaba (China).

Who created DeepSeek-R1-Distill-Qwen?

DeepSeek-R1-Distill-Qwen was created by DeepSeek.

What is the context window for DeepSeek-R1-Distill-Qwen?

A context window is not available for the default version.

Is DeepSeek-R1-Distill-Qwen open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer DeepSeek-R1-Distill-Qwen?

1 provider offerings are linked to the default version.

What models should I compare DeepSeek-R1-Distill-Qwen with?

Nearby ranked alternatives include Mistral Large 3, Gemma 3 27B, Nova 2 Sonic.