← All models
Compare

Gemini 3.1 Pro vs Grok 4.20

Add, remove, or swap models to compare them side by side.

Gemini 3.1 ProGrok 4.20
AttributeGemini 3.1 ProGoogle DeepMind · 3 in the newsGrok 4.20xAI
Scores
Intelligence (ECI)155152
Coding6251
Math7261
Reasoning & Knowledge9456
Agentic & Tools8161
Specifications
DeveloperGoogle DeepMindxAI
FamilyGeminiGrok
ReleasedFeb 19, 2026Feb 17, 2026
Parameters500B
AvailabilityAPI accessAPI access
Context window2M
Price — $/M input$1.25
Price — $/M output$2.50
Inputstext, image, file
Outputstext
Benchmarks
AIME 2024/202596%92%
APEX34%
ARC-AGI98%90%
ARC-AGI-277%65%
FrontierMath37%
FrontierMath Tier 417%
GPQA Diamond94%89%
GSO (code optimization)23%
Humanity's Last Exam46%
METR task horizon6.4 h
SimpleBench80%
SimpleQA Verified77%35%
SWE-bench Verified76%
Terminal-Bench80%57%
WebDev Arena14611395
WeirdML72%52%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.