llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

DeepSeek model product

DeepSeek-V3.1

1 is a hybrid model supporting both thinking and non-thinking modes through different chat templates.

Updated Aug 12, 2026. Default version: DeepSeek-V3.1

Compare
LLMBoard score39.9DeepSeek-V3.1
Coverage80%16 benchmark families
Context window131.1KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

DeepSeek-V3.1 Specifications

Technical details for the model's default version.

Version
DeepSeek-V3.1
Released
Jan 10, 2025
Knowledge cutoff
Unknown
Parameters
671B
Context window
131.1K
Max output
8.2K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

DeepSeek-V3.1 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek-V3.1 category scores

DeepSeek-V3.1 Benchmark Results

Benchmark scores for DeepSeek-V3.1.

16 rows
Columns

Show columns

SimpleQA93.4%034695.6%CAug 11, 2026
Aider-Polyglot68.4%082266.7%CAug 11, 2026
BrowseComp-zh49.2%101325.0%CAug 11, 2026
CodeForces69.7%121626.7%CAug 11, 2026
Terminal-Bench31.3%182529.2%CAug 11, 2026
MMLU-Redux91.8%214857.5%CAug 11, 2026
SWE-bench Multilingual54.5%293415.2%CAug 11, 2026
MMLU-Pro83.7%3012977.3%CAug 11, 2026
HMMT 202533.5%32333.1%CAug 11, 2026
LiveCodeBench56.4%367351.4%CAug 11, 2026
AIME 202466.3%435319.2%CAug 11, 2026
BrowseComp30.0%55585.3%CAug 11, 2026
Humanity's Last Exam15.9%699326.1%CAug 11, 2026
SWE-Bench Verified66.0%7210531.7%CAug 11, 2026
AIME 202549.8%10011412.4%CAug 11, 2026
GPQA74.9%10923453.6%CAug 11, 2026

DeepSeek-V3.1 Arena Results

Preference and agent-evaluation results for the default version.

8 rows
Columns

Show columns

textoverall971419.63,469N/AAug 10, 2026
textoverall1001419.215,027N/AAug 10, 2026
textoverall1071416.53,702N/AAug 10, 2026
textoverall1081416.511,799N/AAug 10, 2026
text style controloverall1131417.215,027N/AAug 10, 2026
text style controloverall1141417.13,469N/AAug 10, 2026
text style controloverall1151416.411,799N/AAug 10, 2026
text style controloverall1191414.53,702N/AAug 10, 2026

DeepSeek-V3.1 Runtime Performance

Provider-specific output speed and catalog latency for DeepSeek-V3.1. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

DeepSeek-V3.1 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.20 input, $0.70 output per 1M via NanoGPT
Tracked offerings
19
18 rows
Columns

Show columns

NanoGPTdeepseek-ai/DeepSeek-V3.1global$0.20$0.70128KAug 11, 2026
submodeldeepseek-ai/DeepSeek-V3.1global$0.20$0.8075KAug 11, 2026
Vercel AI Gatewaydeepseek/deepseek-v3.1global$0.25$0.95163.8KAug 11, 2026
OpenRouterdeepseek/deepseek-chat-v3.1global$0.25$0.95163.8KAug 11, 2026
Deep Infradeepseek-ai/DeepSeek-V3.1global$0.25$0.95163.8KAug 11, 2026
NovitaAIdeepseek/deepseek-v3.1global$0.27$1131.1KAug 11, 2026
Hugging Facedeepseek-ai/DeepSeek-V3.1global$0.27$1131.1KAug 11, 2026
Jiekou.AIdeepseek/deepseek-v3.1global$0.27$1163.8KAug 11, 2026
Meganovadeepseek-ai/DeepSeek-V3.1global$0.27$1164KAug 11, 2026
Basetendeepseek-ai/DeepSeek-V3.1global$0.50$1.5164KAug 11, 2026
Merge Gatewaydeepseek/deepseek-v3.1global$0.50$1.5164KAug 11, 2026
Weights & Biasesdeepseek-ai/DeepSeek-V3.1global$0.55$1.65161KAug 11, 2026
Abacusdeepseek/deepseek-v3.1global$0.55$1.66128KAug 11, 2026
Pioneerdeepseek-ai/DeepSeek-V3.1global$0.56$1.68163.8KAug 11, 2026
Alibaba (China)deepseek-v3-1global$0.574$1.72131.1KAug 11, 2026
Amazon Bedrockdeepseek.v3-v1:0global$0.58$1.68163.8KAug 11, 2026
Together AIdeepseek-ai/DeepSeek-V3-1global$0.60$1.7131.1KAug 11, 2026
Vertexdeepseek-ai/deepseek-v3.1-maasglobal$0.60$1.7163.8KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-V3.1 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

DeepSeek-V3.1Jan 10, 202539.9671B131.1K8.2KYesMIT

What is DeepSeek-V3.1?

Key information about DeepSeek-V3.1 and its available data.

1 is a hybrid model supporting both thinking and non-thinking modes through different chat templates. 1-Base with a two-phase long context extension (32K phase: 630B tokens, 128K phase: 209B tokens), it features 671B total parameters with 37B activated.

Key improvements include smarter tool calling through post-training optimization, higher thinking efficiency achieving comparable quality to DeepSeek-R1-0528 while responding more quickly, and UE8M0 FP8 scale data format for model weights and activations.

The model excels in both reasoning tasks (thinking mode) and practical applications (non-thinking mode), with particularly strong performance in code agent tasks, math competitions, and search-based problem solving.

Data as of 2026-08-11.

DeepSeek-V3.1 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-V3.1vsGrok 4.3DeepSeek-V3.1vsQwen3 MaxDeepSeek-V3.1vsQwen3 VL 32B ThinkingDeepSeek-V3.1vsClaude Sonnet 4DeepSeek-V3.1vsMercury 2DeepSeek-V3.1vsClaude Sonnet 3.7

Models similar to DeepSeek-V3.1

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#111+4.3
DE

DeepSeek-R1

DeepSeek

44.2 LLMBoard

DetailsCompare
#169-13.9
DE

DeepSeek-V3

DeepSeek

26.0 LLMBoard

DetailsCompare
#66+18.2
DE

DeepSeek-V3.2

DeepSeek

58.1 LLMBoard

DetailsCompare
#205-24.2
DE

DeepSeek-R1-Distill-Qwen

DeepSeek

15.7 LLMBoard

DetailsCompare
#206-24.5
DE

DeepSeek-V2.5

DeepSeek

15.4 LLMBoard

DetailsCompare
#215-27.5
DE

DeepSeek-R1-Distill-Llama

DeepSeek

12.4 LLMBoard

DetailsCompare

FAQ

Common questions about DeepSeek-V3.1.

When was DeepSeek-V3.1 released?

DeepSeek-V3.1's default version was released on Jan 10, 2025.

How much does DeepSeek-V3.1 cost?

No official standard PAYG price is currently available for DeepSeek-V3.1. The lowest tracked third-party offer starts at $0.20 input and $0.70 output via NanoGPT.

Who created DeepSeek-V3.1?

DeepSeek-V3.1 was created by DeepSeek.

What is the context window for DeepSeek-V3.1?

The default version has a 131.1K token context window.

Is DeepSeek-V3.1 open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer DeepSeek-V3.1?

19 provider offerings are linked to the default version.

What models should I compare DeepSeek-V3.1 with?

Nearby ranked alternatives include Grok 4.3, Qwen3 Max, Qwen3 VL 32B Thinking.