llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Moonshot AI model product

Kimi K2

Kimi K2 is a Mixture-of-Experts large language model from Moonshot AI with 32 billion activated parameters, 1 trillion total parameters, and a 256K-token context length.

Updated Sep 8, 2026. Default version: Kimi K2-Instruct-0905

LLMBoard Score36.3Kimi K2-Instruct-0905
Coverage100%26 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Kimi K2 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Kimi K2-Instruct-0905 LLMBoard score breakdown

Kimi K2 Benchmark Results

Benchmark scores for Kimi K2-Instruct-0905.

29 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkACEBenchScore76.50%Rank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkAutoLogiScore89.50%Rank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkCNMO 2024Score74.30%Rank02Participants3Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkPolyMath-enScore65.10%Rank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMultiPL-EScore85.70%Rank05Participants13Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkZebraLogicScore89.00%Rank06Participants8Percentile28.57%EvidenceCEvaluatedSep 8, 2026
BenchmarkMATH-500Score97.40%Rank07Participants32Percentile80.65%EvidenceCEvaluatedSep 8, 2026
BenchmarkOJBenchScore27.10%Rank09Participants9Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveBenchScore76.40%Rank10Participants38Percentile75.68%EvidenceCEvaluatedSep 8, 2026
BenchmarkAider-PolyglotScore60.00%Rank13Participants22Percentile42.86%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLUScore89.50%Rank14Participants101Percentile87.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMulti-ChallengeScore54.10%Rank15Participants29Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkIFEvalScore89.80%Rank17Participants68Percentile76.12%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ReduxScore92.70%Rank17Participants48Percentile65.96%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 AirlineScore56.50%Rank18Participants24Percentile26.09%EvidenceCEvaluatedSep 8, 2026
BenchmarkSuperGPQAScore57.20%Rank22Participants34Percentile36.36%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 RetailScore70.60%Rank22Participants27Percentile19.23%EvidenceCEvaluatedSep 8, 2026
BenchmarkTerminal-BenchScore25.00%Rank23Participants25Percentile8.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkSimpleQAScore31.00%Rank25Participants47Percentile47.83%EvidenceCEvaluatedSep 8, 2026
BenchmarkTau2 TelecomScore65.80%Rank29Participants36Percentile20.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkHMMT 2025Score38.80%Rank30Participants33Percentile9.38%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-bench MultilingualScore47.30%Rank37Participants43Percentile14.29%EvidenceCEvaluatedSep 8, 2026
BenchmarkAIME 2024Score69.60%Rank42Participants53Percentile21.15%EvidenceCEvaluatedSep 8, 2026
BenchmarkLiveCodeBenchScore53.70%Rank42Participants75Percentile44.59%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMLU-ProScore81.10%Rank51Participants138Percentile63.50%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-Bench VerifiedScore65.80%Rank78Participants113Percentile31.25%EvidenceCEvaluatedSep 8, 2026
BenchmarkHumanity's Last ExamScore4.70%Rank102Participants103Percentile0.98%EvidenceCEvaluatedSep 8, 2026
BenchmarkAIME 2025Score49.50%Rank108Participants119Percentile9.32%EvidenceCEvaluatedSep 8, 2026
BenchmarkGPQAScore75.10%Rank116Participants247Percentile53.25%EvidenceCEvaluatedSep 8, 2026

Kimi K2 Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Kimi K2 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $1 input, $3 output per 1M via Hugging Face
Tracked offerings
1
1 row
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderHugging FaceProvider model IDmoonshotai/Kimi-K2-Instruct-0905RegionglobalInput / 1M$1Output / 1M$3Context262.1KUpdatedSep 8, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Kimi K2 Runtime Performance

Provider-specific output speed and catalog latency for Kimi K2-Instruct-0905. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Kimi K2 Specifications

Technical details for the model's default version.

Version
Kimi K2-Instruct-0905
Released
Sep 5, 2025
Knowledge cutoff
Unknown
Parameters
1T
Context window
N/A
Max output
N/A
Inputs
text
Outputs
text
Open weights
Yes
License
MIT

Kimi K2 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

4 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionKimi K2 0905ReleasedSep 5, 2025LLMBoardN/AParameters1TContext262.1KMax output262.1KOpen weightsNoLicenseProprietary
VersionKimi K2-Instruct-0905ReleasedSep 5, 2025LLMBoard36.25Parameters1TContextN/AMax outputN/AOpen weightsYesLicenseMIT
VersionKimi K2 BaseReleasedJul 11, 2025LLMBoardN/AParameters1TContextN/AMax outputN/AOpen weightsYesLicenseMIT
VersionKimi K2 InstructReleasedJul 11, 2025LLMBoardN/AParameters1TContext200KMax output200KOpen weightsYesLicenseMIT

Models similar to Kimi K2

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#186-7.23
MA

Kimi k1.5

Moonshot AI

29.02 LLMBoard

Details
#70+23.74
MA

Kimi K2 Thinking

Moonshot AI

59.99 LLMBoard

Details
#59+29.28
MA

Kimi K2.7 Code

Moonshot AI

65.53 LLMBoard

Details
#58+29.32
MA

Kimi K2.5

Moonshot AI

65.57 LLMBoard

Details
#39+37.29
MA

Kimi K2.6

Moonshot AI

73.54 LLMBoard

Details
#8+53.20
MA

Kimi K3

Moonshot AI

89.45 LLMBoard

Details

What is Kimi K2?

Key information about Kimi K2 and its available data.

5T tokens for agentic tasks, coding, mathematics, knowledge, and tool use. It uses the MuonClip optimizer and a hybrid architecture without long thinking.

Data as of 2026-09-08.

FAQ

Common questions about Kimi K2.

When was Kimi K2 released?

Kimi K2's default version was released on Sep 5, 2025.

How much does Kimi K2 cost?

No official standard PAYG price is currently available for Kimi K2. The lowest tracked third-party offer starts at $1 input and $3 output via Hugging Face.

Who created Kimi K2?

Kimi K2 was created by Moonshot AI.

What is the context window for Kimi K2?

A context window is not available for the default version.

Is Kimi K2 open weight?

Yes. The default version is marked as open weight under MIT.

How many API providers offer Kimi K2?

1 provider offerings are linked to the default version.

Browse runtime rankings