← All models
Compare

Claude Opus 4.5 vs Kimi K3

Add, remove, or swap models to compare them side by side.

Claude Opus 4.5Kimi K3
AttributeClaude Opus 4.5AnthropicKimi K3Moonshot · 44 in the news
Scores
Intelligence (ECI)150158
Coding6297
Math2880
Reasoning & Knowledge5573
Agentic & Tools5272
Specifications
DeveloperAnthropicMoonshot
FamilyClaudeKimi
ReleasedNov 24, 2025Jul 16, 2026
Parameters2.8T
AvailabilityAPI accessOpen weights (non-commercial)
Context window200K1M
Price — $/M input$5.00$2.10
Price — $/M output$25.00$10.95
Inputsfile, image, texttext, image, video
Outputstexttext
Benchmarks
AIME 2024/202586%97%
APEX21%51%
ARC-AGI80%95%
ARC-AGI-238%60%
FrontierMath21%
FrontierMath Tier 44%
GDPval (win/tie rate)60%
GPQA Diamond86%93%
GSO (code optimization)27%
Humanity's Last Exam25%47%
LiveCodeBench87%
METR task horizon4.9 h
MMLU-Pro90%
SciCode50%60%
SimpleBench62%61%
SimpleQA Verified46%51%
SWE-bench Verified77%
Terminal-Bench63%
WebDev Arena14941674
WeirdML64%83%
τ²-bench89%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.