← All models
Compare

Claude Sonnet 5.5 (batch) vs Kimi K3

Add, remove, or swap models to compare them side by side.

Claude Sonnet 5.5 (batch)Kimi K3
AttributeClaude Sonnet 5.5 (batch)AnthropicKimi K3Moonshot · 46 in the news
Scores
Intelligence (ECI)—158
Coding—94
Math—74
Reasoning & Knowledge—70
Agentic & Tools—67
Specifications
DeveloperAnthropicMoonshot
FamilyClaudeKimi
ReleasedSep 28, 2026Jul 16, 2026
Parameters—2.8T
Availability—Open weights (non-commercial)
Context window1M1M
Price — $/M input$1.00$1.99
Price — $/M output$5.00$9.00
Inputstext, image, filetext, image, video
Outputstexttext
Benchmarks
AIME 2024/2025—97%
APEX—51%
ARC-AGI—95%
ARC-AGI-2—60%
GPQA Diamond—93%
Humanity's Last Exam—47%
SciCode—60%
SimpleBench—61%
SimpleQA Verified—51%
WebDev Arena—1674
WeirdML—83%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.