← All models
Compare

Claude Opus 5.5 vs Grok 4.6

Claude Opus 5.5 (Anthropic) and Grok 4.6 (xAI) compared on benchmarks, pricing, context window, and use-case rankings.

AttributeClaude Opus 5.5Anthropic · 3 in the newsGrok 4.6xAI · 3 in the news
Scores
Intelligence (ECI)156
Coding10089
Math91
Reasoning & Knowledge10073
Agentic & Tools95
Specifications
DeveloperAnthropicxAI
FamilyClaudeGrok
ReleasedSep 22, 2026Aug 12, 2026
Parameters
AvailabilityAPI access
Context window1M500K
Price — $/M input$4.00$2.00
Price — $/M output$20.00$6.00
Inputstext, image, filetext, image, file
Outputstexttext
Benchmarks
Humanity's Last Exam61%43%
SciCode67%56%
AIME 2024/202599%
APEX65%
ARC-AGI88%
ARC-AGI-267%
GPQA Diamond94%
SimpleBench76%
SimpleQA Verified49%
WebDev Arena1618
WeirdML67%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Frequently asked questions

Which is cheaper, Claude Opus 5.5 or Grok 4.6?

Grok 4.6 is cheaper on input tokens at $2.00 per million, versus $4.00 (representative OpenRouter pricing).

Which has a larger context window, Claude Opus 5.5 or Grok 4.6?

Claude Opus 5.5 supports up to 1M tokens, compared with 500K for the other.

Which is better for coding, Claude Opus 5.5 or Grok 4.6?

Across coding benchmarks like SWE-bench Verified and Terminal-Bench, Claude Opus 5.5 ranks higher — 100th vs 89th percentile among the models tracked on Model Beat.

Want a different match-up? Open the compare tool to add or swap models.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.