Kimi K2.7 Code

Kimi K2.7 Code is a reasoning model from Moonshot AI. 34 benchmarks count toward its score, in 6 categories.

availableShows if the model has enough results for an index.
IndexOverall score out of 100.59.5 ±4.3
CoverageShare of the index weight with results.85%
SpeedOutput tokens per second.58/s
Input / 1MUS dollars per 1M input tokens.$0.706
Output / 1MUS dollars per 1M output tokens.$3.21
ContextMaximum tokens in one request.262K
EloLMArena rating and rank.N/A

The index is a score out of 100. The ± range shows how much it can change.

CapabilitiesScore per category, out of 100.

Out of 100
AgenticMulti-step tasks with tools.
60.6
CodingCode writing and repair.
60.7
ReasoningLogic problems and puzzles.
57.5
MultimodalTasks with images and text.
N/A
KnowledgeFacts and expert knowledge.
60.0
MultilingualTasks in many languages.
N/A
InstructionTasks with strict rules in the prompt.
51.1
MathMath problems.
58.7

Results

34 counted
BenchmarkThe test name.CategoryThe capability that the test measures.ResultThe score from the publisher.IndexThis result as a score out of 100.RunThe settings of the run.DateDate of the result.Published byThe source of the result.
OTIS Mock AIME 2024-2025Math95.6%65.0Epoch AI
τ²-Bench Tool-Agent-User EvaluationAgentic90.1%63.5Victor Barres et al.
Artificial Analysis GPQA DiamondKnowledge89.6%60.4Artificial Analysis
GPQA diamondKnowledge87.9%59.4Epoch AI
LiveBench ReasoningReasoning82.8%71.025 Jun 2026LiveBench
LiveCodeBenchCoding82.0%58.91 Sept 2026Vals AI
MCPMark-VerifiedAgentic81.1%MCPMark
LiveBench MathematicsMath79.6%54.225 Jun 2026LiveBench
Artificial Analysis Long Context ReasoningReasoning79.3%63.1Artificial Analysis
SWE-benchCoding78.2%60.31 Sept 2026Vals AI
LiveBench LanguageKnowledge77.9%64.625 Jun 2026LiveBench
MCP AtlasAgentic76.0%65.4OpenAI
LiveBench CodingCoding74.0%60.525 Jun 2026LiveBench
Terminal-Bench 2.1Agentic67.0%63.621 Sept 2026Vals AI
Artificial Analysis IFBenchInstruction63.1%53.1Artificial Analysis
LiveBench Data AnalysisReasoning62.7%43.025 Jun 2026LiveBench
Kimi Code Bench v2Coding62.0%Moonshot AI
Artificial Analysis Coding IndexCoding60.8%61.8Artificial Analysis
LiveBench Instruction FollowingInstruction56.3%49.125 Jun 2026LiveBench
FrontierMath-Tiers-1-3-v2-PrivateMath54.0%61.9Epoch AI
ProgramBench: Can Language Models Rebuild Programs From Scratch?Coding53.6%58.5John Yang et al.
OpenHarmony Bench v1.0Coding52.1%60.9OpenHarmony Bench authors
SkillsBenchCoding50.0%65.4OpenHands11 Sept 2026Vals AI
cursorBench32Coding49.7%58.9Benchmark authors
Artificial Analysis SciCodeCoding47.8%59.2Artificial Analysis
Vibe Code Bench v1.1Coding47.2%61.8OpenHands21 Sept 2026Vals AI
Kimi Claw 24/7 BenchAgentic46.9%Moonshot AI
LiveBench Agentic CodingAgentic45.7%59.725 Jun 2026LiveBench
Artificial Analysis Omniscience AccuracyKnowledge39.6%62.9Artificial Analysis
SimpleQA VerifiedKnowledge36.5%55.3Epoch AI
MLS-Bench LiteCoding35.1%MLS-Bench
Artificial Analysis Humanity's Last ExamKnowledge35.0%62.5Artificial Analysis
GDPval-AA normalizedAgentic26.3%55.8Artificial Analysis
Artificial Analysis Intelligence IndexKnowledge25.8%54.7Artificial Analysis
Code MigrationCoding25.4%61.221 Sept 2026Vals AI
Artificial Analysis Agentic IndexAgentic22.5%55.5Artificial Analysis
Chess PuzzlesReasoning21.0%50.7Epoch AI
FrontierMath-Tier-4-v2-PrivateMath12.2%53.8Epoch AI
Critical Physics TasksReasoning10.0%59.3Artificial Analysis
ProgramBenchCoding0.0%21 Sept 2026Vals AI

34 benchmarks count, from 35 of 40 results. A grey row does not count. Too few models took that benchmark.

Sources

BenchLM benchmark aggregationUsed with attribution; per-benchmark results credited to their original authorsOpenRouter, collected directlyNo licence statedEpoch AI, collected directlyCC BY — free to use and redistribute with attributionLiveBench, collected directlyNo licence stated for the leaderboard. The site repo serving these CSVs has no LICENSE; the harness repo carries upstream Apache-2.0 and MIT copies that cover code, not resultsVals AI, collected directlyNo licence stated. Read from the public leaderboard and credited to Vals AI

Same level, lower price

Qwen3.8-Flash-Next66.4 · Freedots3-note Preview65.5 · FreeApodex 1.1 Mini59.1 · Free

More from Moonshot AI

Kimi K372.0Kimi K2.660.9Kimi K2.5 (Reasoning)54.6Kimi K2.553.3Kimi K2.5 Thinking53.8