← All models
Compare

Claude Haiku 4.5 vs Gemini 3.1 Pro

Claude Haiku 4.5 (Anthropic) and Gemini 3.1 Pro (Google DeepMind) compared on benchmarks, pricing, context window, and use-case rankings.

AttributeClaude Haiku 4.5Anthropic · 1 in the newsGemini 3.1 ProGoogle DeepMind · 3 in the news
Scores
Intelligence (ECI)143155
Coding2565
Math1374
Reasoning & Knowledge1294
Agentic & Tools2383
Specifications
DeveloperAnthropicGoogle DeepMind
FamilyClaudeGemini
ReleasedOct 15, 2025Feb 19, 2026
Parameters
AvailabilityAPI accessAPI access
Context window200K
Price — $/M input$1.00
Price — $/M output$5.00
Inputstext, image, file
Outputstext
Benchmarks
AIME 2024/202567%96%
APEX9%34%
ARC-AGI48%98%
ARC-AGI-24%77%
FrontierMath6%37%
FrontierMath Tier 42%17%
GPQA Diamond71%94%
MATH Level 596%
SimpleQA Verified6%77%
Terminal-Bench36%80%
WebDev Arena13221461
WeirdML45%72%
GSO (code optimization)23%
Humanity's Last Exam46%
METR task horizon6.4 h
SimpleBench80%
SWE-bench Verified76%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Frequently asked questions

Is Claude Haiku 4.5 better than Gemini 3.1 Pro?

On Epoch AI's Capabilities Index, Gemini 3.1 Pro scores higher (155) than Claude Haiku 4.5 (143). The right pick depends on your task — compare their coding, math, and reasoning scores in the table above.

Which is better for coding, Claude Haiku 4.5 or Gemini 3.1 Pro?

Across coding benchmarks like SWE-bench Verified and Terminal-Bench, Gemini 3.1 Pro ranks higher — 65th vs 25th percentile among the models tracked on Model Beat.

Want a different match-up? Open the compare tool to add or swap models.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.