llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Moonshot AI model product

Kimi K2

Kimi K2 0905 is the September update of Kimi K2 0711.

Updated Aug 17, 2026. Default version: Kimi K2-Instruct-0905

Compare
LLMBoard Score36.6Kimi K2-Instruct-0905
Coverage100%26 benchmark families
Context windowN/ATokens
Official input price$0.60Moonshot AI API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Kimi K2 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Kimi K2-Instruct-0905 LLMBoard score breakdown

Kimi K2 Benchmark Results

Benchmark scores for Kimi K2-Instruct-0905.

29 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkACEBenchScore76.5%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAutoLogiScore89.5%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCNMO 2024Score74.3%Rank02Participants3Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkPolyMath-enScore65.1%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMultiPL-EScore85.7%Rank05Participants13Percentile66.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkZebraLogicScore89.0%Rank06Participants8Percentile28.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMATH-500Score97.4%Rank07Participants32Percentile80.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkOJBenchScore27.1%Rank09Participants9Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveBenchScore76.4%Rank10Participants38Percentile75.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkAider-PolyglotScore60.0%Rank13Participants22Percentile42.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLUScore89.5%Rank14Participants101Percentile87.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMulti-ChallengeScore54.1%Rank15Participants29Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkIFEvalScore89.8%Rank17Participants67Percentile75.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ReduxScore92.7%Rank17Participants48Percentile66.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTau2 AirlineScore56.5%Rank17Participants23Percentile27.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkTau2 RetailScore70.6%Rank21Participants26Percentile20.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSuperGPQAScore57.2%Rank22Participants34Percentile36.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-BenchScore25.0%Rank23Participants25Percentile8.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkSimpleQAScore31.0%Rank25Participants47Percentile47.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkTau2 TelecomScore65.8%Rank28Participants35Percentile20.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkHMMT 2025Score38.8%Rank30Participants33Percentile9.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-bench MultilingualScore47.3%Rank34Participants38Percentile10.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2024Score69.6%Rank42Participants53Percentile21.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBenchScore53.7%Rank42Participants75Percentile44.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProScore81.1%Rank50Participants134Percentile63.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore65.8%Rank78Participants111Percentile30.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore4.7%Rank98Participants99Percentile1.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2025Score49.5%Rank104Participants115Percentile9.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore75.1%Rank111Participants239Percentile53.8%EvidenceCEvaluatedAug 17, 2026

Kimi K2 Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Kimi K2 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.60 input, $2.5 output per 1M
Official provider
Moonshot AI
Lowest third-party
From $1 input, $3 output per 1M via Hugging Face
Tracked offerings
2
1 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderHugging FaceProvider model IDmoonshotai/Kimi-K2-Instruct-0905RegionglobalInput / 1M$1Output / 1M$3Context262.1KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Kimi K2 Runtime Performance

Provider-specific output speed and catalog latency for Kimi K2-Instruct-0905. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Kimi K2 Specifications

Technical details for the model's default version.

Version
Kimi K2-Instruct-0905
Released
Sep 5, 2025
Knowledge cutoff
Unknown
Parameters
1T
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
No
License
MIT

Kimi K2 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

4 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionKimi K2 0905ReleasedSep 5, 2025LLMBoardN/AParameters1TContext262.1KMax output262.1KOpen weightsNoLicenseProprietary
VersionKimi K2-Instruct-0905ReleasedSep 5, 2025LLMBoard36.6Parameters1TContextN/AMax outputN/AOpen weightsNoLicenseMIT
VersionKimi K2 BaseReleasedJul 11, 2025LLMBoardN/AParameters1TContextN/AMax outputN/AOpen weightsNoLicenseMIT
VersionKimi K2 InstructReleasedJul 11, 2025LLMBoardN/AParameters1TContext200KMax output200KOpen weightsNoLicenseMIT

Kimi K2 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Kimi K2vsClaude Sonnet 3.7Kimi K2vsGemma 4 12BKimi K2vsNemotron 3.5 LightningKimi K2vsNorth Mini Code 1.0Kimi K2vsMiniMax M1 80KKimi K2vsLongCat Flash Chat

Models similar to Kimi K2

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#173-7.1
MA

Kimi k1.5

Moonshot AI

29.5 LLMBoard

DetailsCompare
#61+24.0
MA

Kimi K2 Thinking

Moonshot AI

60.6 LLMBoard

DetailsCompare
#52+29.1
MA

Kimi K2.5

Moonshot AI

65.7 LLMBoard

DetailsCompare
#49+30.4
MA

Kimi K2.7 Code

Moonshot AI

67.0 LLMBoard

DetailsCompare
#30+38.6
MA

Kimi K2.6

Moonshot AI

75.1 LLMBoard

DetailsCompare
#5+55.4
MA

Kimi K3

Moonshot AI

92.0 LLMBoard

DetailsCompare

What is Kimi K2?

Key information about Kimi K2 and its available data.

Kimi K2 0905 is the September update of Kimi K2 0711. It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It supports long-context inference up to 256k tokens, extended from the previous 128k.

This update improves agentic coding with higher accuracy and better generalization across scaffolds, and enhances frontend coding with more aesthetic and functional outputs for web, 3D, and related tasks. The model is trained with a novel stack incorporating the MuonClip optimizer for stable large-scale MoE training.

Data as of 2026-08-17.

FAQ

Common questions about Kimi K2.

When was Kimi K2 released?

Kimi K2's default version was released on Sep 5, 2025.

How much does Kimi K2 cost?

Kimi K2's official API price is $0.60 per million input tokens and $2.5 per million output tokens via Moonshot AI. The lowest tracked third-party offer starts at $1 input and $3 output via Hugging Face.

Who created Kimi K2?

Kimi K2 was created by Moonshot AI.

What is the context window for Kimi K2?

A context window is not available for the default version.

Is Kimi K2 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Kimi K2?

2 provider offerings are linked to the default version.

What models should I compare Kimi K2 with?

Nearby ranked alternatives include Claude Sonnet 3.7, Gemma 4 12B, Nemotron 3.5 Lightning.