← All models
Compare

Gemini 3.1 Pro vs Grok 4.6

Add, remove, or swap models to compare them side by side.

Gemini 3.1 ProGrok 4.6
AttributeGemini 3.1 ProGoogle DeepMind · 3 in the newsGrok 4.6xAI · 1 in the news
Scores
Intelligence (ECI)155156
Coding5991
Math7193
Reasoning & Knowledge9376
Agentic & Tools7989
Specifications
DeveloperGoogle DeepMindxAI
FamilyGeminiGrok
ReleasedFeb 19, 2026Aug 12, 2026
Parameters
AvailabilityAPI accessAPI access
Context window500K
Price — $/M input$2.00
Price — $/M output$6.00
Inputstext, image, file
Outputstext
Benchmarks
AIME 2024/202596%99%
APEX34%41%
ARC-AGI98%88%
ARC-AGI-277%67%
FrontierMath37%
FrontierMath Tier 417%
GPQA Diamond94%94%
GSO (code optimization)23%
Humanity's Last Exam46%43%
METR task horizon6.4 h
SimpleBench80%
SimpleQA Verified77%54%
SWE-bench Verified76%
Terminal-Bench80%
WebDev Arena14611631
WeirdML72%67%
SciCode54%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.