← All models
Compare

Claude Sonnet 5.5 (batch) vs Kimi K3

Claude Sonnet 5.5 (batch) (Anthropic) and Kimi K3 (Moonshot) compared on benchmarks, pricing, context window, and use-case rankings.

AttributeClaude Sonnet 5.5 (batch)AnthropicKimi K3Moonshot · 46 in the news
Scores
Intelligence (ECI)—158
Coding—94
Math—74
Reasoning & Knowledge—70
Agentic & Tools—67
Specifications
DeveloperAnthropicMoonshot
FamilyClaudeKimi
ReleasedSep 28, 2026Jul 16, 2026
Parameters—2.8T
Availability—Open weights (non-commercial)
Context window1M1M
Price — $/M input$1.00$1.99
Price — $/M output$5.00$9.00
Inputstext, image, filetext, image, video
Outputstexttext
Benchmarks
AIME 2024/2025—97%
APEX—51%
ARC-AGI—95%
ARC-AGI-2—60%
GPQA Diamond—93%
Humanity's Last Exam—47%
SciCode—60%
SimpleBench—61%
SimpleQA Verified—51%
WebDev Arena—1674
WeirdML—83%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Frequently asked questions

Which is cheaper, Claude Sonnet 5.5 (batch) or Kimi K3?

Claude Sonnet 5.5 (batch) is cheaper on input tokens at $1.00 per million, versus $1.99 (representative OpenRouter pricing).

Which has a larger context window, Claude Sonnet 5.5 (batch) or Kimi K3?

Kimi K3 supports up to 1M tokens, compared with 1M for the other.

Want a different match-up? Open the compare tool to add or swap models.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.