MoonshotAI: Kimi K2 Thinking.

See how MoonshotAI: Kimi K2 Thinking compares with other broadly useful AI models across quality, speed, price, context, and practical capabilities.

CompareVisit model

Quality

#31/67

17.3

Overall quality score

Speed

#31/67

79 t/s

Output tokens per second

Input price

#40/67

$0.60/1M

Per 1M tokens

Output price

#41/67

$2.50/1M

Per 1M tokens

Context

#37/67

262.1K

Content read at once

Comparison summary

MoonshotAI: Kimi K2 Thinking ranks #31 of 67 for overall quality and #31 for response speed in the current catalogue.

Its listed input price is $0.60/1M, output price is $2.50/1M, and it can work with up to 262.1K of context at once.

Its strongest areas in the current data are coding, everyday help, math and logic.

Practical specifications

Company
Moonshot AI
Reasoning
Yes
Open weights
Yes
Input
Text
Output
Text
Context window
262.1K

MoonshotAI: Kimi K2 Thinking technical benchmarks.

Named evaluation records for technical comparison. Results retain their original benchmark names and sources.

BenchmarkAreaKimi K2 ThinkingSource
Intelligence IndexoverallQuality17.3OpenRouter / artificial-analysis
Coding Indexcoding21OpenRouter / artificial-analysis
Agentic Indexagentic1.8OpenRouter / artificial-analysis
GPQAGPQA Diamond71.3OpenRouter model benchmarks
Humanity's Last ExamHLE9.5OpenRouter model benchmarks
IFBenchIFBench62.8OpenRouter model benchmarks
τ²-Bench Telecomτ²-Bench Telecom25.4OpenRouter model benchmarks
AA-LCRAA-LCR52.7OpenRouter model benchmarks
GDPval-AAGDPval-AA0OpenRouter model benchmarks
CritPtCritPt0OpenRouter model benchmarks
SciCodeSciCode33OpenRouter model benchmarks
Terminal-Bench HardTerminal-Bench Hard6.8OpenRouter model benchmarks
AA-Omniscience AccuracyAA-Omniscience Accuracy15.7OpenRouter model benchmarks
AA-Omniscience Non-Hallucination RateAA-Omniscience Non-Hallucination Rate41.1OpenRouter model benchmarks
MMLU-ProKnowledge and reasoning84.6OpenEvals/leaderboard-data
SWE-bench VerifiedCoding71.3OpenEvals/leaderboard-data
Terminal-BenchAgentic coding35.7OpenEvals/leaderboard-data

Quality benchmarks

See where MoonshotAI: Kimi K2 Thinking stands.

The darker bar marks MoonshotAI: Kimi K2 Thinking; the lighter bars provide context from other leading models.

Highlights

Quality

Overall capability · Higher is better

Speed

Output tokens per second · Higher is better

Input price

USD per 1M tokens · Lower is better