llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Moonshot AI model product

Kimi K3

Kimi K3 is Moonshot AI’s open Mixture-of-Experts language model for coding, knowledge work, and reasoning.

Updated Sep 8, 2026. Default version: Kimi K3

LLMBoard Score89.5Kimi K3
Coverage20%26 benchmark families
Context window1MTokens
Official input price$3Moonshot AI API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Kimi K3 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Kimi K3 LLMBoard score breakdown

Kimi K3 Benchmark Results

Benchmark scores for Kimi K3.

30 of 41 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkBabyVisionScore85.70%Rank01Participants11Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkDECK-BenchScore73.50%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkDeepSearchQAScore95.00%Rank01Participants10Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkKimi Code Bench v2Score72.90%Rank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMathVisionScore97.80%Rank01Participants35Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkMLS-Bench LiteScore48.30%Rank01Participants3Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkOmniDocBenchScore91.10%Rank01Participants2Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkProgram BenchScore77.80%Rank01Participants7Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSpreadsheetBench 2Score34.80%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkZEROBenchScore0.41 pointsRank01Participants10Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkAA-BriefcaseScore1,548.00 pointsRank02Participants3Percentile50.00%EvidenceBEvaluatedSep 8, 2026
BenchmarkAPEX-AgentsScore37.60%Rank02Participants9Percentile87.50%EvidenceCEvaluatedSep 8, 2026
BenchmarkBrowseCompScore91.20%Rank02Participants63Percentile98.39%EvidenceCEvaluatedSep 8, 2026
BenchmarkCharXiv-RScore91.30%Rank02Participants55Percentile98.15%EvidenceCEvaluatedSep 8, 2026
BenchmarkFrontierSWEScore81.20%Rank02Participants16Percentile93.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkMCP AtlasScore84.20%Rank02Participants34Percentile96.97%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMMU-Pro (with tools)Score83.40%Rank02Participants4Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkPerceptionBenchScore58.50%Rank02Participants2Percentile0.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkSWE-MarathonScore42.00%Rank02Participants5Percentile75.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Task Outcome ExplicitScore16.58%Rank02Participants49Percentile97.92%EvidenceAEvaluatedSep 5, 2026
BenchmarkDeepSWEScore67.50%Rank03Participants13Percentile83.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkPostTrainBenchScore36.60%Rank03Participants7Percentile66.67%EvidenceCEvaluatedSep 8, 2026
BenchmarkWorldVQAScore51.00%Rank03Participants5Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkTerminal-Bench 2.1Score88.30%Rank04Participants35Percentile91.18%EvidenceCEvaluatedSep 8, 2026
BenchmarkOfficeQA ProScore63.30%Rank05Participants9Percentile50.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena WebdevScore1,674.35 ratingRank05Participants97Percentile95.83%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent LeaderboardScore7.96%Rank06Participants49Percentile89.58%EvidenceAEvaluatedSep 5, 2026
BenchmarkDeepSWE 1.1Score69.00%Rank07Participants31Percentile80.00%EvidenceBEvaluatedSep 8, 2026
BenchmarkJob BenchScore52.90%Rank07Participants8Percentile14.29%EvidenceCEvaluatedSep 8, 2026
BenchmarkMMMU-ProScore81.60%Rank07Participants69Percentile91.18%EvidenceCEvaluatedSep 8, 2026

Kimi K3 Arena Results

Preference and agent-evaluation results for the default version.

30 of 94 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenaagent task outcome explicitCategoryoverallRank02Rating / score0.17VotesN/AObservations74.2KResult dateSep 5, 2026
Arenatext style controlCategoryjapaneseRank03Rating / score1,515.57Votes319ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorykoreanRank03Rating / score1,478.00Votes425ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryjapaneseRank04Rating / score1,502.29Votes319ObservationsN/AResult dateSep 2, 2026
ArenatextCategorykoreanRank04Rating / score1,465.75Votes425ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorycodingRank04Rating / score1,542.01Votes4,719ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry legal and governmentRank04Rating / score1,509.44Votes1,437ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorypolishRank04Rating / score1,502.78Votes292ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry mathematicalRank05Rating / score1,488.75Votes704ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryindustry medicine and healthcareRank05Rating / score1,513.53Votes1,011ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryhard promptsRank05Rating / score1,518.56Votes11,751ObservationsN/AResult dateSep 2, 2026
ArenawebdevCategoryoverallRank05Rating / score1,674.35Votes4,546ObservationsN/AResult dateSep 5, 2026
ArenawebdevCategorywebdevRank05Rating / score1,674.35Votes4,546ObservationsN/AResult dateSep 5, 2026
ArenawebdevCategorywebdev-reactRank05Rating / score1,692.50Votes3,279ObservationsN/AResult dateSep 5, 2026
ArenaagentCategoryoverallRank06Rating / score0.08VotesN/AObservations8.2MResult dateSep 5, 2026
ArenatextCategoryindustry legal and governmentRank06Rating / score1,502.69Votes1,437ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry mathematicalRank06Rating / score1,514.81Votes898ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry software and it servicesRank06Rating / score1,528.22Votes7,085ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymathRank06Rating / score1,505.73Votes760ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymulti turnRank06Rating / score1,499.18Votes2,887ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryspanishRank06Rating / score1,488.46Votes555ObservationsN/AResult dateSep 2, 2026
Arenaagent praise complaintCategoryoverallRank07Rating / score0.14VotesN/AObservations28.8KResult dateSep 5, 2026
ArenatextCategorycodingRank07Rating / score1,510.68Votes4,719ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryhard promptsRank07Rating / score1,497.47Votes11,751ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry mathematicalRank07Rating / score1,509.12Votes898ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry software and it servicesRank07Rating / score1,503.04Votes7,085ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryspanishRank07Rating / score1,480.36Votes555ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryexpertRank07Rating / score1,528.02Votes1,619ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryhard prompts englishRank07Rating / score1,517.41Votes4,559ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry life and physical and social scienceRank07Rating / score1,517.97Votes2,837ObservationsN/AResult dateSep 2, 2026

Kimi K3 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$3 input, $15 output per 1M
Official provider
Moonshot AI
Lowest third-party
From $2 input, $10 output per 1M via NanoGPT
Tracked offerings
62
30 of 61 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderUmans AI Coding PlanProvider model IDumans-kimi-k3RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderNvidiaProvider model IDmoonshotai/kimi-k3RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderKenariProvider model IDkimi-k3RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderSCNet Token PlanProvider model IDKimi-K3RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderKimi For CodingProvider model IDk3RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderSenseNova (China)Provider model IDkimi-k3RegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDmoonshotai/kimi-k3RegionglobalInput / 1M$2Output / 1M$10Context1MUpdatedSep 8, 2026
ProviderCrofAIProvider model IDkimi-k3RegionglobalInput / 1M$2Output / 1M$8Context1MUpdatedSep 8, 2026
ProviderRequestyProvider model IDkimi-k3RegionglobalInput / 1M$2.25Output / 1M$11.25Context1MUpdatedSep 8, 2026
ProviderVancineProvider model IDkimi-k3RegionglobalInput / 1M$2.4Output / 1M$12Context1MUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDkimi-k3RegionglobalInput / 1M$2.83Output / 1M$14.13Context1MUpdatedSep 8, 2026
ProviderDeep InfraProvider model IDmoonshotai/Kimi-K3RegionglobalInput / 1M$2.85Output / 1M$14.25Context1MUpdatedSep 8, 2026
ProviderDigitalOceanProvider model IDkimi-k3RegionglobalInput / 1M$2.85Output / 1M$14.25Context1MUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDmoonshot/kimi-k3RegionglobalInput / 1M$2.9Output / 1M$14Context1MUpdatedSep 8, 2026
ProviderNeuralwattProvider model IDkimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderTokenGoProvider model IDmoonshotai/kimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderCortecsProvider model IDkimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderSyntheticProvider model IDhf:moonshotai/Kimi-K3RegionglobalInput / 1M$3Output / 1M$15Context524.3KUpdatedSep 8, 2026
ProviderJalapeno CloudProvider model IDKimi-K3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderHugging FaceProvider model IDmoonshotai/Kimi-K3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderImpossiblProvider model IDmoonshotai/kimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDmoonshotai/kimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderMoonshot AIProvider model IDkimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderCrossModelProvider model IDmoonshot/kimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderOpenCode GoProvider model IDkimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderOpperProvider model IDmoonshot/kimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderVivgridProvider model IDkimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderNebius Token FactoryProvider model IDmoonshotai/Kimi-K3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderCloudflare AI GatewayProvider model IDmoonshotai/kimi-k3RegionglobalInput / 1M$3Output / 1M$15Context1MUpdatedSep 8, 2026
ProviderBerget.AIProvider model IDmoonshotai/Kimi-K3RegionglobalInput / 1M$3Output / 1M$15Context327.7KUpdatedSep 8, 2026

Kimi K3 Runtime Performance

Provider-specific output speed and catalog latency for Kimi K3. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Kimi K3 Specifications

Technical details for the model's default version.

Version
Kimi K3
Released
Jul 16, 2026
Knowledge cutoff
Unknown
Parameters
2.8T
Context window
1M
Max output
1M
Inputs
image, text, video
Outputs
text
Open weights
Yes
License
Kimi K3 License

Kimi K3 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionKimi K3ReleasedJul 16, 2026LLMBoard89.45Parameters2.8TContext1MMax output1MOpen weightsYesLicenseKimi K3 License

Models similar to Kimi K3

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#39-15.91
MA

Kimi K2.6

Moonshot AI

73.54 LLMBoard

Details
#58-23.88
MA

Kimi K2.5

Moonshot AI

65.57 LLMBoard

Details
#59-23.92
MA

Kimi K2.7 Code

Moonshot AI

65.53 LLMBoard

Details
#70-29.46
MA

Kimi K2 Thinking

Moonshot AI

59.99 LLMBoard

Details
#158-53.20
MA

Kimi K2

Moonshot AI

36.25 LLMBoard

Details
#186-60.43
MA

Kimi k1.5

Moonshot AI

29.02 LLMBoard

Details

What is Kimi K3?

Key information about Kimi K3 and its available data.

8 trillion total parameters, 896 experts, and 16 active experts. It supports a 1M-token context window, native image and video understanding, tool calling, structured output, and always-on reasoning, and combines Kimi Delta Attention, Attention Residuals, and Stable LatentMoE.

Data as of 2026-09-08.

FAQ

Common questions about Kimi K3.

When was Kimi K3 released?

Kimi K3's default version was released on Jul 16, 2026.

How much does Kimi K3 cost?

Kimi K3's official API price is $3 per million input tokens and $15 per million output tokens via Moonshot AI. The lowest tracked third-party offer starts at $2 input and $10 output via NanoGPT.

Who created Kimi K3?

Kimi K3 was created by Moonshot AI.

What is the context window for Kimi K3?

The default version has a 1M token context window.

Is Kimi K3 open weight?

Yes. The default version is marked as open weight under Kimi K3 License.

How many API providers offer Kimi K3?

62 provider offerings are linked to the default version.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Browse runtime rankings