llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

DeepSeek model product

DeepSeek-V3.1

1 is a hybrid model supporting both thinking and non-thinking modes through different chat templates.

Updated Aug 17, 2026. Default version: DeepSeek-V3.1

Compare
LLMBoard Score39.2DeepSeek-V3.1
Coverage80%16 benchmark families
Context window131.1KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

DeepSeek-V3.1 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek-V3.1 LLMBoard score breakdown

DeepSeek-V3.1 Benchmark Results

Benchmark scores for DeepSeek-V3.1.

16 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkSimpleQAScore93.4%Rank03Participants47Percentile95.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkAider-PolyglotScore68.4%Rank08Participants22Percentile66.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseComp-zhScore49.2%Rank10Participants13Percentile25.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCodeForcesScore69.7%Rank13Participants17Percentile25.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-BenchScore31.3%Rank18Participants25Percentile29.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ReduxScore91.8%Rank21Participants48Percentile57.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkHMMT 2025Score33.5%Rank32Participants33Percentile3.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-bench MultilingualScore54.5%Rank32Participants38Percentile16.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProScore83.7%Rank33Participants134Percentile75.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBenchScore56.4%Rank38Participants75Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2024Score66.3%Rank43Participants53Percentile19.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseCompScore30.0%Rank59Participants62Percentile4.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore15.9%Rank74Participants99Percentile25.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore66.0%Rank77Participants111Percentile30.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2025Score49.8%Rank101Participants115Percentile12.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore74.9%Rank114Participants239Percentile52.5%EvidenceCEvaluatedAug 17, 2026

DeepSeek-V3.1 Arena Results

Preference and agent-evaluation results for the default version.

8 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
ArenatextCategoryoverallRank97Rating / score1419.8Votes3,468ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank101Rating / score1419.1Votes15,022ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank108Rating / score1416.5Votes3,702ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank109Rating / score1416.4Votes11,796ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank113Rating / score1417.3Votes3,468ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank114Rating / score1417.2Votes15,022ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank116Rating / score1416.3Votes11,796ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank120Rating / score1414.5Votes3,702ObservationsN/AResult dateAug 12, 2026

DeepSeek-V3.1 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.20 input, $0.70 output per 1M via NanoGPT
Tracked offerings
19
18 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderNanoGPTProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.20Output / 1M$0.70Context128KUpdatedAug 17, 2026
ProvidersubmodelProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.20Output / 1M$0.80Context75KUpdatedAug 17, 2026
ProviderDeep InfraProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.25Output / 1M$0.95Context163.8KUpdatedAug 17, 2026
ProviderOpenRouterProvider model IDdeepseek/deepseek-chat-v3.1RegionglobalInput / 1M$0.25Output / 1M$0.95Context163.8KUpdatedAug 17, 2026
ProviderVercel AI GatewayProvider model IDdeepseek/deepseek-v3.1RegionglobalInput / 1M$0.25Output / 1M$0.95Context163.8KUpdatedAug 17, 2026
ProviderMeganovaProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.27Output / 1M$1Context164KUpdatedAug 17, 2026
ProviderHugging FaceProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.27Output / 1M$1Context131.1KUpdatedAug 17, 2026
ProviderJiekou.AIProvider model IDdeepseek/deepseek-v3.1RegionglobalInput / 1M$0.27Output / 1M$1Context163.8KUpdatedAug 17, 2026
ProviderNovitaAIProvider model IDdeepseek/deepseek-v3.1RegionglobalInput / 1M$0.27Output / 1M$1Context131.1KUpdatedAug 17, 2026
ProviderMerge GatewayProvider model IDdeepseek/deepseek-v3.1RegionglobalInput / 1M$0.50Output / 1M$1.5Context164KUpdatedAug 17, 2026
ProviderBasetenProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.50Output / 1M$1.5Context164KUpdatedAug 17, 2026
ProviderWeights & BiasesProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.55Output / 1M$1.65Context161KUpdatedAug 17, 2026
ProviderAbacusProvider model IDdeepseek/deepseek-v3.1RegionglobalInput / 1M$0.55Output / 1M$1.66Context128KUpdatedAug 17, 2026
ProviderPioneerProvider model IDdeepseek-ai/DeepSeek-V3.1RegionglobalInput / 1M$0.56Output / 1M$1.68Context163.8KUpdatedAug 17, 2026
ProviderAlibaba (China)Provider model IDdeepseek-v3-1RegionglobalInput / 1M$0.574Output / 1M$1.72Context131.1KUpdatedAug 17, 2026
ProviderAmazon BedrockProvider model IDdeepseek.v3-v1:0RegionglobalInput / 1M$0.58Output / 1M$1.68Context163.8KUpdatedAug 17, 2026
ProviderVertexProvider model IDdeepseek-ai/deepseek-v3.1-maasRegionglobalInput / 1M$0.60Output / 1M$1.7Context163.8KUpdatedAug 17, 2026
ProviderTogether AIProvider model IDdeepseek-ai/DeepSeek-V3-1RegionglobalInput / 1M$0.60Output / 1M$1.7Context131.1KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-V3.1 Runtime Performance

Provider-specific output speed and catalog latency for DeepSeek-V3.1. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

DeepSeek-V3.1 Specifications

Technical details for the model's default version.

Version
DeepSeek-V3.1
Released
Jan 10, 2025
Knowledge cutoff
Unknown
Parameters
671B
Context window
131.1K
Max output
8.2K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

DeepSeek-V3.1 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionDeepSeek-V3.1ReleasedJan 10, 2025LLMBoard39.2Parameters671BContext131.1KMax output8.2KOpen weightsYesLicenseMIT

DeepSeek-V3.1 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-V3.1vsQwen3 MaxDeepSeek-V3.1vsQwen3 VL 32B ThinkingDeepSeek-V3.1vsGPT-5.3DeepSeek-V3.1vsMercury 2DeepSeek-V3.1vsClaude Sonnet 4DeepSeek-V3.1vsGemini 2.5 Flash

Models similar to DeepSeek-V3.1

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#122+4.3
DE

DeepSeek-R1

DeepSeek

43.5 LLMBoard

DetailsCompare
#181-13.7
DE

DeepSeek-V3

DeepSeek

25.5 LLMBoard

DetailsCompare
#74+18.3
DE

DeepSeek-V3.2

DeepSeek

57.5 LLMBoard

DetailsCompare
#217-24.4
DE

DeepSeek-R1-Distill-Qwen

DeepSeek

14.8 LLMBoard

DetailsCompare
#221-25.4
DE

DeepSeek-V2.5

DeepSeek

13.8 LLMBoard

DetailsCompare
#227-27.6
DE

DeepSeek-R1-Distill-Llama

DeepSeek

11.6 LLMBoard

DetailsCompare

What is DeepSeek-V3.1?

Key information about DeepSeek-V3.1 and its available data.

1 is a hybrid model supporting both thinking and non-thinking modes through different chat templates. 1-Base with a two-phase long context extension (32K phase: 630B tokens, 128K phase: 209B tokens), it features 671B total parameters with 37B activated.

Key improvements include smarter tool calling through post-training optimization, higher thinking efficiency achieving comparable quality to DeepSeek-R1-0528 while responding more quickly, and UE8M0 FP8 scale data format for model weights and activations.

The model excels in both reasoning tasks (thinking mode) and practical applications (non-thinking mode), with particularly strong performance in code agent tasks, math competitions, and search-based problem solving.

Data as of 2026-08-17.

FAQ

Common questions about DeepSeek-V3.1.

When was DeepSeek-V3.1 released?

DeepSeek-V3.1's default version was released on Jan 10, 2025.

How much does DeepSeek-V3.1 cost?

No official standard PAYG price is currently available for DeepSeek-V3.1. The lowest tracked third-party offer starts at $0.20 input and $0.70 output via NanoGPT.

Who created DeepSeek-V3.1?

DeepSeek-V3.1 was created by DeepSeek.

What is the context window for DeepSeek-V3.1?

The default version has a 131.1K token context window.

Is DeepSeek-V3.1 open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer DeepSeek-V3.1?

19 provider offerings are linked to the default version.

What models should I compare DeepSeek-V3.1 with?

Nearby ranked alternatives include Qwen3 Max, Qwen3 VL 32B Thinking, GPT-5.3.