llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Moonshot AI model product

Kimi K2 Thinking

Kimi K2 Thinking is the latest, most capable version of open-source thinking model.

Updated Aug 17, 2026. Default version: Kimi K2-Thinking-0905

Compare
LLMBoard Score60.6Kimi K2-Thinking-0905
Coverage100%19 benchmark families
Context window262.1KTokens
Official input price$0.60Moonshot AI API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Kimi K2 Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Kimi K2-Thinking-0905 LLMBoard score breakdown

Kimi K2 Thinking Benchmark Results

Benchmark scores for Kimi K2-Thinking-0905.

21 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkFinSearchComp-T3Score47.4%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkFRAMESScore87.0%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHealthBenchScore58.0%Rank02Participants9Percentile87.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkOJBenchScore48.7%Rank02Participants9Percentile87.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkSeal-0Score56.3%Rank02Participants6Percentile80.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-BenchScore47.1%Rank03Participants25Percentile91.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkHMMT 2025Score97.5%Rank04Participants33Percentile90.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMulti-SWE-BenchScore41.9%Rank04Participants6Percentile40.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2025Score100.0%Rank05Participants115Percentile96.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ReduxScore94.4%Rank05Participants48Percentile91.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseComp-zhScore62.3%Rank08Participants13Percentile41.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkSciCodeScore44.8%Rank09Participants21Percentile60.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBench v6Score83.1%Rank15Participants56Percentile74.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkWritingBenchScore73.8%Rank15Participants15Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkIMO-AnswerBenchScore78.6%Rank18Participants20Percentile10.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore51.0%Rank20Participants99Percentile80.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProScore84.6%Rank27Participants134Percentile80.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-bench MultilingualScore61.1%Rank28Participants38Percentile27.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseCompScore60.2%Rank37Participants62Percentile41.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore71.3%Rank58Participants111Percentile48.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore84.5%Rank62Participants239Percentile74.4%EvidenceCEvaluatedAug 17, 2026

Kimi K2 Thinking Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Kimi K2 Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.60 input, $2.5 output per 1M
Official provider
Moonshot AI
Lowest third-party
From $0.40 input, $2.5 output per 1M via OpenCode Zen
Tracked offerings
6
6 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderOpenCode ZenProvider model IDkimi-k2-thinkingRegionglobalInput / 1M$0.40Output / 1M$2.5Context262.1KUpdatedAug 17, 2026
ProviderHeliconeProvider model IDkimi-k2-thinkingRegionglobalInput / 1M$0.48Output / 1M$2Context256KUpdatedAug 17, 2026
Provider302.AIProvider model IDkimi-k2-thinkingRegionglobalInput / 1M$0.575Output / 1M$2.3Context262.1KUpdatedAug 17, 2026
ProviderMoonshot AIProvider model IDkimi-k2-thinkingRegionglobalInput / 1M$0.60Output / 1M$2.5Context262.1KUpdatedAug 17, 2026
ProviderLLM GatewayProvider model IDkimi-k2-thinkingRegionglobalInput / 1M$0.60Output / 1M$2.5Context262.1KUpdatedAug 17, 2026
ProviderMoonshot AI (China)Provider model IDkimi-k2-thinkingRegionglobalInput / 1M$0.60Output / 1M$2.5Context262.1KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Kimi K2 Thinking Runtime Performance

Provider-specific output speed and catalog latency for Kimi K2-Thinking-0905. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Kimi K2 Thinking Specifications

Technical details for the model's default version.

Version
Kimi K2-Thinking-0905
Released
Sep 5, 2025
Knowledge cutoff
Unknown
Parameters
1T
Context window
262.1K
Max output
262.1K
Inputs
text
Outputs
text
Open weights
No
License
MIT

Kimi K2 Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionKimi K2-Thinking-0905ReleasedSep 5, 2025LLMBoard60.6Parameters1TContext262.1KMax output262.1KOpen weightsNoLicenseMIT

Kimi K2 Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Kimi K2 ThinkingvsGPT-5.1Kimi K2 ThinkingvsMiMoKimi K2 ThinkingvsGPT-5.1-InstantKimi K2 ThinkingvsStep 3.5 FlashKimi K2 ThinkingvsLongCat Flash ThinkingKimi K2 ThinkingvsGPT-5.4-mini

Models similar to Kimi K2 Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#52+5.1
MA

Kimi K2.5

Moonshot AI

65.7 LLMBoard

DetailsCompare
#49+6.4
MA

Kimi K2.7 Code

Moonshot AI

67.0 LLMBoard

DetailsCompare
#30+14.6
MA

Kimi K2.6

Moonshot AI

75.1 LLMBoard

DetailsCompare
#146-24.0
MA

Kimi K2

Moonshot AI

36.6 LLMBoard

DetailsCompare
#173-31.1
MA

Kimi k1.5

Moonshot AI

29.5 LLMBoard

DetailsCompare
#5+31.5
MA

Kimi K3

Moonshot AI

92.0 LLMBoard

DetailsCompare

What is Kimi K2 Thinking?

Key information about Kimi K2 Thinking and its available data.

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, it is built as a thinking agent that reasons step-by-step while dynamically invoking tools.

It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks by dramatically scaling multi-step reasoning depth and maintaining stable tool-use across 200–300 sequential calls.

At the same time, K2 Thinking is a native INT4 quantization model with 256k context window, achieving lossless reductions in inference latency and GPU memory usage.

Key features include deep thinking & tool orchestration with end-to-end training to interleave chain-of-thought reasoning with function calls, native INT4 quantization via Quantization-Aware Training (QAT) achieving lossless 2x speed-up, and stable long-horizon agency maintaining coherent goal-directed behavior across up to 200–300 consecutive tool invocations.

Data as of 2026-08-17.

FAQ

Common questions about Kimi K2 Thinking.

When was Kimi K2 Thinking released?

Kimi K2 Thinking's default version was released on Sep 5, 2025.

How much does Kimi K2 Thinking cost?

Kimi K2 Thinking's official API price is $0.60 per million input tokens and $2.5 per million output tokens via Moonshot AI. The lowest tracked third-party offer starts at $0.40 input and $2.5 output via OpenCode Zen.

Who created Kimi K2 Thinking?

Kimi K2 Thinking was created by Moonshot AI.

What is the context window for Kimi K2 Thinking?

The default version has a 262.1K token context window.

Is Kimi K2 Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Kimi K2 Thinking?

6 provider offerings are linked to the default version.

What models should I compare Kimi K2 Thinking with?

Nearby ranked alternatives include GPT-5.1, MiMo, GPT-5.1-Instant.