llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Moonshot AI model product

Kimi K2 Thinking

Kimi K2 Thinking is the latest, most capable version of open-source thinking model.

Updated Aug 12, 2026. Default version: Kimi K2-Thinking-0905

Compare
LLMBoard score61.0Kimi K2-Thinking-0905
Coverage100%19 benchmark families
Context window262.1KTokens
Official input price$0.60Moonshot AI API

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Kimi K2 Thinking Specifications

Technical details for the model's default version.

Version
Kimi K2-Thinking-0905
Released
Sep 5, 2025
Knowledge cutoff
Unknown
Parameters
1T
Context window
262.1K
Max output
262.1K
Inputs
text
Outputs
text
Open weights
No
License
MIT

Kimi K2 Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Kimi K2-Thinking-0905 category scores

Kimi K2 Thinking Benchmark Results

Benchmark scores for Kimi K2-Thinking-0905.

21 rows
Columns

Show columns

FinSearchComp-T347.4%011100.0%CAug 11, 2026
FRAMES87.0%012100.0%CAug 11, 2026
HealthBench58.0%02987.5%CAug 11, 2026
OJBench48.7%02987.5%CAug 11, 2026
Seal-056.3%02680.0%CAug 11, 2026
Terminal-Bench47.1%032591.7%CAug 11, 2026
HMMT 202597.5%043390.6%CAug 11, 2026
Multi-SWE-Bench41.9%04640.0%CAug 11, 2026
AIME 2025100.0%0511496.5%CAug 11, 2026
MMLU-Redux94.4%054891.5%CAug 11, 2026
BrowseComp-zh62.3%081341.7%CAug 11, 2026
SciCode44.8%081961.1%CAug 11, 2026
LiveCodeBench v683.1%135376.9%CAug 11, 2026
WritingBench73.8%15150.0%CAug 11, 2026
IMO-AnswerBench78.6%171911.1%CAug 11, 2026
Humanity's Last Exam51.0%189381.5%CAug 11, 2026
MMLU-Pro84.6%2412982.0%CAug 11, 2026
SWE-bench Multilingual61.1%253427.3%CAug 11, 2026
BrowseComp60.2%365838.6%CAug 11, 2026
SWE-Bench Verified71.3%5510548.1%CAug 11, 2026
GPQA84.5%5823475.5%CAug 11, 2026

Kimi K2 Thinking Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Kimi K2 Thinking Runtime Performance

Provider-specific output speed and catalog latency for Kimi K2-Thinking-0905. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Kimi K2 Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.60 input, $2.5 output per 1M
Official provider
Moonshot AI
Lowest third-party
From $0.40 input, $2.5 output per 1M via OpenCode Zen
Tracked offerings
6
6 rows
Columns

Show columns

OpenCode Zenkimi-k2-thinkingglobal$0.40$2.5262.1KAug 11, 2026
Heliconekimi-k2-thinkingglobal$0.48$2256KAug 11, 2026
302.AIkimi-k2-thinkingglobal$0.575$2.3262.1KAug 11, 2026
LLM Gatewaykimi-k2-thinkingglobal$0.60$2.5262.1KAug 11, 2026
Moonshot AI (China)kimi-k2-thinkingglobal$0.60$2.5262.1KAug 11, 2026
Moonshot AIkimi-k2-thinkingglobal$0.60$2.5262.1KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Kimi K2 Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Kimi K2-Thinking-0905Sep 5, 202561.01T262.1K262.1KNoMIT

What is Kimi K2 Thinking?

Key information about Kimi K2 Thinking and its available data.

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, it is built as a thinking agent that reasons step-by-step while dynamically invoking tools.

It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks by dramatically scaling multi-step reasoning depth and maintaining stable tool-use across 200–300 sequential calls.

At the same time, K2 Thinking is a native INT4 quantization model with 256k context window, achieving lossless reductions in inference latency and GPU memory usage.

Key features include deep thinking & tool orchestration with end-to-end training to interleave chain-of-thought reasoning with function calls, native INT4 quantization via Quantization-Aware Training (QAT) achieving lossless 2x speed-up, and stable long-horizon agency maintaining coherent goal-directed behavior across up to 200–300 consecutive tool invocations.

Data as of 2026-08-11.

Kimi K2 Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Kimi K2 ThinkingvsGPT-5.1Kimi K2 ThinkingvsMiMoKimi K2 ThinkingvsGPT-5.1-InstantKimi K2 ThinkingvsLongCat Flash ThinkingKimi K2 ThinkingvsQwen3.6 27BKimi K2 ThinkingvsGPT-5.4-mini

Models similar to Kimi K2 Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#46+5.0
MA

Kimi K2.5

Moonshot AI

66.0 LLMBoard

DetailsCompare
#44+5.9
MA

Kimi K2.7 Code

Moonshot AI

66.8 LLMBoard

DetailsCompare
#26+14.5
MA

Kimi K2.6

Moonshot AI

75.5 LLMBoard

DetailsCompare
#136-23.8
MA

Kimi K2

Moonshot AI

37.2 LLMBoard

DetailsCompare
#162-30.8
MA

Kimi k1.5

Moonshot AI

30.1 LLMBoard

DetailsCompare
#5+32.1
MA

Kimi K3

Moonshot AI

93.0 LLMBoard

DetailsCompare

FAQ

Common questions about Kimi K2 Thinking.

When was Kimi K2 Thinking released?

Kimi K2 Thinking's default version was released on Sep 5, 2025.

How much does Kimi K2 Thinking cost?

Kimi K2 Thinking's official API price is $0.60 per million input tokens and $2.5 per million output tokens via Moonshot AI. The lowest tracked third-party offer starts at $0.40 input and $2.5 output via OpenCode Zen.

Who created Kimi K2 Thinking?

Kimi K2 Thinking was created by Moonshot AI.

What is the context window for Kimi K2 Thinking?

The default version has a 262.1K token context window.

Is Kimi K2 Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Kimi K2 Thinking?

6 provider offerings are linked to the default version.

What models should I compare Kimi K2 Thinking with?

Nearby ranked alternatives include GPT-5.1, MiMo, GPT-5.1-Instant.