llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

DeepSeek model product

DeepSeek-V4-Flash

DeepSeek-V4-Flash-0423 is the preview release of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window, evaluated here at the default high reasoning effort.

Updated Aug 17, 2026. Default version: DeepSeek-V4-Flash-0731

Compare
LLMBoard Score77.4DeepSeek-V4-Flash-0731
Coverage20%7 benchmark families
Context window1MTokens
Official input price$0.14DeepSeek API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

DeepSeek-V4-Flash Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

DeepSeek-V4-Flash-0731 LLMBoard score breakdown

DeepSeek-V4-Flash Benchmark Results

Benchmark scores for DeepSeek-V4-Flash-0731.

9 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkDSBench-FullStackScore68.7%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkDSBench-HardScore59.6%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkNL2RepoScore54.2%Rank04Participants17Percentile81.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkDeepSWEScore54.4%Rank06Participants11Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkToolathlonScore70.3%Rank06Participants37Percentile86.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkAutomationBenchScore25.1%Rank07Participants12Percentile45.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkCyberGymScore76.7%Rank07Participants13Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAgents' Last ExamScore25.2%Rank10Participants10Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-Bench 2.1Score82.7%Rank12Participants28Percentile59.3%EvidenceCEvaluatedAug 17, 2026

DeepSeek-V4-Flash Arena Results

Preference and agent-evaluation results for the default version.

11 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenaagent task outcome explicitCategoryoverallRank10Rating / score0.1VotesN/AObservations21.3KResult dateAug 13, 2026
ArenawebdevCategoryoverallRank11Rating / score1581.2Votes2,936ObservationsN/AResult dateAug 14, 2026
Arenaagent tool hallucinationCategoryoverallRank14Rating / score0.0VotesN/AObservations3MResult dateAug 13, 2026
ArenaagentCategoryoverallRank19Rating / score0.0VotesN/AObservations3.2MResult dateAug 13, 2026
Arenaagent steerabilityCategoryoverallRank20Rating / score0.0VotesN/AObservations30.1KResult dateAug 13, 2026
Arenaagent praise complaintCategoryoverallRank23Rating / score0.0VotesN/AObservations8KResult dateAug 13, 2026
Arenaagent bash recovery stepsCategoryoverallRank27Rating / score0.0VotesN/AObservations110.2KResult dateAug 13, 2026
ArenawebdevCategoryoverallRank56Rating / score1431.1Votes9,493ObservationsN/AResult dateAug 14, 2026
Arenatext style controlCategoryoverallRank79Rating / score1438.3Votes48,789ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank88Rating / score1424.0Votes48,789ObservationsN/AResult dateAug 12, 2026
Arenatext factualityCategoryoverallRank88Rating / score1433.9Votes48,702ObservationsN/AResult dateAug 12, 2026

DeepSeek-V4-Flash Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.14 input, $0.28 output per 1M
Official provider
DeepSeek
Lowest third-party
From $0.08 input, $0.18 output per 1M via Deep Infra
Tracked offerings
46
30 of 45 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderUmans AI Coding PlanProvider model IDumans-deepseek-v4-flash-0731RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderAlibaba Token PlanProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderHetznerProvider model IDDeepSeek-V4-Flash-0731RegionglobalInput / 1MN/AOutput / 1MN/AContext512KUpdatedAug 17, 2026
ProviderAlibaba Token Plan (China)Provider model IDdeepseek-v4-flash-0731RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderDeep InfraProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.08Output / 1M$0.18Context1MUpdatedAug 17, 2026
ProviderDigitalOceanProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.08Output / 1M$0.252Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDdeepinfra/deepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.08Output / 1M$0.18Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDflexai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.08Output / 1M$0.18Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDflexai/deepseek-v4-flash-0731RegionglobalInput / 1M$0.08Output / 1M$0.18Context1MUpdatedAug 17, 2026
ProviderCrofAIProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.12Output / 1M$0.21Context1MUpdatedAug 17, 2026
ProviderCortecsProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderInceptronProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderPerplexity AgentProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.26Context1MUpdatedAug 17, 2026
ProviderBasetenProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.26Context1MUpdatedAug 17, 2026
ProviderVercel AI GatewayProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.26Context1MUpdatedAug 17, 2026
ProviderRunInfraProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.27Context1MUpdatedAug 17, 2026
ProviderWeights & BiasesProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.13Output / 1M$0.28Context262.1KUpdatedAug 17, 2026
ProviderNanoGPTProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderDeepSeekProvider model IDdeepseek-v4-flashRegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderHugging FaceProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderOpenRouterProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1.3MUpdatedAug 17, 2026
ProviderAmbientProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderMerge GatewayProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderUmans AIProvider model IDumans-deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderTogether AIProvider model IDdeepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderOpenCode ZenProvider model IDdeepseek-v4-flashRegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderRequestyProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderEmpirioLabs AIProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderAMDProvider model IDDeepSeek-V4-FlashRegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDtogether_ai/deepseek-ai/DeepSeek-V4-Flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDfireworks_ai/accounts/fireworks/models/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderKilo GatewayProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderFireworks AIProvider model IDaccounts/fireworks/models/deepseek-v4-flash-0731RegionglobalInput / 1M$0.14Output / 1M$0.28Context1MUpdatedAug 17, 2026
ProviderVivgridProvider model IDdeepseek-v4-flashRegionglobalInput / 1M$0.15Output / 1M$0.30Context1MUpdatedAug 17, 2026
ProviderGreenPTProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.1596Output / 1M$0.399Context1MUpdatedAug 17, 2026
ProviderVenice AIProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.175Output / 1M$0.35Context1MUpdatedAug 17, 2026
ProviderAlibabaProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.20Output / 1M$0.40Context1MUpdatedAug 17, 2026
ProviderCharm HyperProvider model IDdeepseek-v4-flash-0731RegionglobalInput / 1M$0.20Output / 1M$0.40Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDqwen/deepseek-v4-flash-0731RegionglobalInput / 1M$0.20Output / 1M$0.40Context1MUpdatedAug 17, 2026
ProviderOpenCode GoProvider model IDdeepseek-v4-flashRegionglobalInput / 1M$0.22Output / 1M$0.66Context1MUpdatedAug 17, 2026
ProviderTensorXProvider model IDdeepseek/deepseek-v4-flash-0731RegionglobalInput / 1M$0.25Output / 1M$0.30Context1MUpdatedAug 17, 2026
ProviderCloudflare Workers AIProvider model ID@cf/deepseek-ai/deepseek-v4-flash-0731RegionglobalInput / 1M$0.44Output / 1M$1.32Context1.3MUpdatedAug 17, 2026
ProviderEden AIProvider model IDcloudflare/@cf/deepseek-ai/deepseek-v4-flash-0731RegionglobalInput / 1M$0.44Output / 1M$1.32Context1MUpdatedAug 17, 2026
ProviderEden AIProvider model IDscaleway/deepseek-v4-flash-0731RegionglobalInput / 1M$0.4627Output / 1M$0.9254Context256KUpdatedAug 17, 2026
ProviderTinfoilProvider model IDdeepseek-v4-flashRegionglobalInput / 1M$0.70Output / 1M$1.9Context1MUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

DeepSeek-V4-Flash Runtime Performance

Provider-specific output speed and catalog latency for DeepSeek-V4-Flash-0731. Runtime does not affect the capability score.

3 rows
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderNovitaOutput Speed19.447 tok/sCatalog Latency4.368 sMax Input1MMax Output393.2KUpdatedAug 17, 2026
ProviderDeepInfraOutput Speed5.85 tok/sCatalog Latency17.843 sMax Input1MMax Output65.5KUpdatedAug 17, 2026
ProviderFireworksOutput Speed0.879 tok/sCatalog Latency8.101 sMax Input1MMax Output65.5KUpdatedAug 17, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

DeepSeek-V4-Flash Specifications

Technical details for the model's default version.

Version
DeepSeek-V4-Flash-0731
Released
Jul 31, 2026
Knowledge cutoff
May 1, 2025
Parameters
304B
Context window
1M
Max output
384K
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

DeepSeek-V4-Flash Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

4 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionDeepSeek-V4-Flash-0731ReleasedJul 31, 2026LLMBoard77.4Parameters304BContext1MMax output384KOpen weightsYesLicenseMIT
VersionDeepSeek V4 FlashReleasedApr 24, 2026LLMBoardN/AParametersN/AContext1MMax output384KOpen weightsYesLicenseUnspecified
VersionDeepSeek-V4-Flash-0423ReleasedApr 23, 2026LLMBoardN/AParameters284BContext1MMax output131.1KOpen weightsNoLicenseMIT
VersionDeepSeek-V4-Flash-MaxReleasedApr 23, 2026LLMBoardN/AParameters284BContext1MMax output393.2KOpen weightsNoLicenseMIT

DeepSeek-V4-Flash vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

DeepSeek-V4-FlashvsQwen3.7 MaxDeepSeek-V4-FlashvsQwen3.8 27BDeepSeek-V4-FlashvsGemini 3.1 ProDeepSeek-V4-FlashvsGPT-5.6-LunaDeepSeek-V4-FlashvsGPT-5.4DeepSeek-V4-FlashvsGPT-5.5-Pro

Models similar to DeepSeek-V4-Flash

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#7+12.6
DE

DeepSeek-V4-Pro

DeepSeek

90.0 LLMBoard

DetailsCompare
#74-19.9
DE

DeepSeek-V3.2

DeepSeek

57.5 LLMBoard

DetailsCompare
#122-33.9
DE

DeepSeek-R1

DeepSeek

43.5 LLMBoard

DetailsCompare
#138-38.2
DE

DeepSeek-V3.1

DeepSeek

39.2 LLMBoard

DetailsCompare
#181-51.9
DE

DeepSeek-V3

DeepSeek

25.5 LLMBoard

DetailsCompare
#217-62.5
DE

DeepSeek-R1-Distill-Qwen

DeepSeek

14.8 LLMBoard

DetailsCompare

What is DeepSeek-V4-Flash?

Key information about DeepSeek-V4-Flash and its available data.

DeepSeek-V4-Flash-0423 is the preview release of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window, evaluated here at the default high reasoning effort.

It shares the V4 series' hybrid attention architecture combining Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA) for dramatically improved long-context efficiency, Manifold-Constrained Hyper-Connections (mHC) for stable signal propagation, and the Muon optimizer for faster convergence.

Pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation, V4-Flash offers reasoning capabilities that closely approach V4-Pro with faster responses and highly cost-effective pricing.

Data as of 2026-08-17.

FAQ

Common questions about DeepSeek-V4-Flash.

When was DeepSeek-V4-Flash released?

DeepSeek-V4-Flash's default version was released on Jul 31, 2026.

How much does DeepSeek-V4-Flash cost?

DeepSeek-V4-Flash's official API price is $0.14 per million input tokens and $0.28 per million output tokens via DeepSeek. The lowest tracked third-party offer starts at $0.08 input and $0.18 output via Deep Infra.

Who created DeepSeek-V4-Flash?

DeepSeek-V4-Flash was created by DeepSeek.

What is the context window for DeepSeek-V4-Flash?

The default version has a 1M token context window.

Is DeepSeek-V4-Flash open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer DeepSeek-V4-Flash?

46 provider offerings are linked to the default version.

What models should I compare DeepSeek-V4-Flash with?

Nearby ranked alternatives include Qwen3.7 Max, Qwen3.8 27B, Gemini 3.1 Pro.