← All models
Compare

DeepSeek-V4-Pro-0813 vs Grok 4.20

DeepSeek-V4-Pro-0813 (DeepSeek) and Grok 4.20 (xAI) compared on benchmarks, pricing, context window, and use-case rankings.

AttributeDeepSeek-V4-Pro-0813DeepSeek · 1 in the newsGrok 4.20xAI
Scores
Intelligence (ECI)155152
Coding7642
Math8858
Reasoning & Knowledge6947
Agentic & Tools6755
Specifications
DeveloperDeepSeekxAI
FamilyDeepSeekGrok
ReleasedAug 13, 2026Feb 17, 2026
Parameters1.6T500B
AvailabilityOpen weights (unrestricted)API access
Context window1M2M
Price — $/M input$0.58$1.25
Price — $/M output$1.74$2.50
Inputstexttext, image, file
Outputstexttext
Benchmarks
AIME 2024/202599%92%
APEX47%
ARC-AGI91%90%
ARC-AGI-261%65%
GPQA Diamond92%89%
Humanity's Last Exam41%
SciCode51%
SimpleQA Verified53%30%
WebDev Arena15811374
WeirdML66%52%
Terminal-Bench57%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Frequently asked questions

Is DeepSeek-V4-Pro-0813 better than Grok 4.20?

On Epoch AI's Capabilities Index, DeepSeek-V4-Pro-0813 scores higher (155) than Grok 4.20 (152). The right pick depends on your task — compare their coding, math, and reasoning scores in the table above.

Which is cheaper, DeepSeek-V4-Pro-0813 or Grok 4.20?

DeepSeek-V4-Pro-0813 is cheaper on input tokens at $0.58 per million, versus $1.25 (representative OpenRouter pricing).

Which has a larger context window, DeepSeek-V4-Pro-0813 or Grok 4.20?

Grok 4.20 supports up to 2M tokens, compared with 1M for the other.

Which is better for coding, DeepSeek-V4-Pro-0813 or Grok 4.20?

Across coding benchmarks like SWE-bench Verified and Terminal-Bench, DeepSeek-V4-Pro-0813 ranks higher — 76th vs 42th percentile among the models tracked on Model Beat.

Want a different match-up? Open the compare tool to add or swap models.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.