← All models
Compare

DeepSeek-V4-Pro-0813 vs Kimi K2 Thinking

Add, remove, or swap models to compare them side by side.

DeepSeek-V4-Pro-0813Kimi K2 Thinking
AttributeDeepSeek-V4-Pro-0813DeepSeek · 1 in the newsKimi K2 ThinkingMoonshot · 6 in the news
Scores
Intelligence (ECI)155146
Coding7637
Math8818
Reasoning & Knowledge6932
Agentic & Tools6723
Specifications
DeveloperDeepSeekMoonshot
FamilyDeepSeekKimi
ReleasedAug 13, 2026Nov 6, 2025
Parameters1.6T1T
AvailabilityOpen weights (unrestricted)Open weights (restricted use)
Context window1M262K
Price — $/M input$0.58$0.60
Price — $/M output$1.74$2.50
Inputstexttext
Outputstexttext
Benchmarks
AIME 2024/202599%83%
APEX47%4%
ARC-AGI91%
ARC-AGI-261%
GPQA Diamond92%84%
Humanity's Last Exam41%24%
SciCode51%42%
SimpleQA Verified53%32%
WebDev Arena15811337
WeirdML66%43%
FrontierMath21%
FrontierMath Tier 40%
LiveCodeBench85%
METR task horizon54 min
MMLU-Pro85%
Terminal-Bench36%
τ²-bench93%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.