← All models
Compare

Claude Sonnet 4.5 vs Kimi K3

Add, remove, or swap models to compare them side by side.

Claude Sonnet 4.5Kimi K3
AttributeClaude Sonnet 4.5Anthropic · 4 in the newsKimi K3Moonshot · 44 in the news
Scores
Intelligence (ECI)147158
Coding2597
Math4180
Reasoning & Knowledge2273
Agentic & Tools1772
Specifications
DeveloperAnthropicMoonshot
FamilyClaudeKimi
ReleasedSep 29, 2025Jul 16, 2026
Parameters2.8T
AvailabilityAPI accessOpen weights (non-commercial)
Context window1M1M
Price — $/M input$3.00$2.10
Price — $/M output$15.00$10.95
Inputstext, image, filetext, image, video
Outputstexttext
Benchmarks
AIME 2024/202578%97%
ARC-AGI64%95%
ARC-AGI-214%60%
FrontierMath15%
FrontierMath Tier 44%
GDPval (win/tie rate)50%
GPQA Diamond82%93%
GSO (code optimization)15%
Humanity's Last Exam14%47%
MATH Level 598%
METR task horizon2.0 h
SimpleBench54%61%
SimpleQA Verified31%51%
SWE-bench Verified71%
Terminal-Bench47%
WebDev Arena13911674
WeirdML48%83%
APEX51%
SciCode60%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.