← All models
Compare

Grok 4.5 vs Kimi K2 Thinking

Add, remove, or swap models to compare them side by side.

Grok 4.5Kimi K2 Thinking
AttributeGrok 4.5xAI · 2 in the newsKimi K2 ThinkingMoonshot · 6 in the news
Scores
Intelligence (ECI)154146
Coding8945
Math8419
Reasoning & Knowledge6838
Agentic & Tools7426
Specifications
DeveloperxAIMoonshot
FamilyGrokKimi
ReleasedJul 8, 2026Nov 6, 2025
Parameters1T
AvailabilityAPI accessOpen weights (restricted use)
Context window500K262K
Price — $/M input$2.00$0.60
Price — $/M output$6.00$2.50
Inputstext, image, filetext
Outputstexttext
Benchmarks
AIME 2024/202598%83%
APEX34%4%
ARC-AGI87%
ARC-AGI-253%
GPQA Diamond93%84%
Humanity's Last Exam40%22%
SciCode54%42%
SimpleBench70%
SimpleQA Verified54%32%
WebDev Arena15511337
WeirdML46%43%
FrontierMath21%
FrontierMath Tier 40%
LiveCodeBench85%
METR task horizon54 min
MMLU-Pro85%
Terminal-Bench36%
τ²-bench93%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.