← All models
Compare

Gemini 3.1 Pro vs Kimi K2 Thinking

Add, remove, or swap models to compare them side by side.

Gemini 3.1 ProKimi K2 Thinking
AttributeGemini 3.1 ProGoogle DeepMind · 3 in the newsKimi K2 ThinkingMoonshot · 6 in the news
Scores
Intelligence (ECI)155146
Coding6245
Math7219
Reasoning & Knowledge9438
Agentic & Tools8126
Specifications
DeveloperGoogle DeepMindMoonshot
FamilyGeminiKimi
ReleasedFeb 19, 2026Nov 6, 2025
Parameters1T
AvailabilityAPI accessOpen weights (restricted use)
Context window262K
Price — $/M input$0.60
Price — $/M output$2.50
Inputstext
Outputstext
Benchmarks
AIME 2024/202596%83%
APEX34%4%
ARC-AGI98%
ARC-AGI-277%
FrontierMath37%21%
FrontierMath Tier 417%0%
GPQA Diamond94%84%
GSO (code optimization)23%
Humanity's Last Exam46%22%
METR task horizon6.4 h54 min
SimpleBench80%
SimpleQA Verified77%32%
SWE-bench Verified76%
Terminal-Bench80%36%
WebDev Arena14611337
WeirdML72%43%
LiveCodeBench85%
MMLU-Pro85%
SciCode42%
τ²-bench93%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.