← All models
Compare

Gemini 3.7 Flash vs gpt-oss-120b

Gemini 3.7 Flash (Google DeepMind) and gpt-oss-120b (OpenAI) compared on benchmarks, pricing, context window, and use-case rankings.

AttributeGemini 3.7 FlashGoogle DeepMind · 1 in the newsgpt-oss-120bOpenAI · 2 in the news
Scores
Intelligence (ECI)157140
Coding9135
Math8042
Reasoning & Knowledge8515
Agentic & Tools987
Specifications
DeveloperGoogle DeepMindOpenAI
FamilyGeminiGPT
ReleasedAug 13, 2026Aug 5, 2025
Parameters116.8B
AvailabilityAPI accessOpen weights (unrestricted)
Context window1M131K
Price — $/M input$0.75$0.03
Price — $/M output$3.75$0.17
Inputstext, image, video, file, audiotext
Outputstexttext
Benchmarks
AIME 2024/202597%89%
APEX68%4%
ARC-AGI96%
ARC-AGI-285%
GPQA Diamond95%76%
Humanity's Last Exam39%20%
SciCode60%34%
SimpleQA Verified69%14%
WebDev Arena1587
LiveCodeBench88%
METR task horizon42 min
MMLU-Pro81%
SimpleBench22%
Terminal-Bench19%
WeirdML48%
τ²-bench66%

Use-case scores are 0–100 percentile composites across each area’s benchmarks, ranked against every model from the past year. Highlighted cells lead each row. Open a model for the full picture.

Frequently asked questions

Is Gemini 3.7 Flash better than gpt-oss-120b?

On Epoch AI's Capabilities Index, Gemini 3.7 Flash scores higher (157) than gpt-oss-120b (140). The right pick depends on your task — compare their coding, math, and reasoning scores in the table above.

Which is cheaper, Gemini 3.7 Flash or gpt-oss-120b?

gpt-oss-120b is cheaper on input tokens at $0.03 per million, versus $0.75 (representative OpenRouter pricing).

Which has a larger context window, Gemini 3.7 Flash or gpt-oss-120b?

Gemini 3.7 Flash supports up to 1M tokens, compared with 131K for the other.

Which is better for coding, Gemini 3.7 Flash or gpt-oss-120b?

Across coding benchmarks like SWE-bench Verified and Terminal-Bench, Gemini 3.7 Flash ranks higher — 91th vs 35th percentile among the models tracked on Model Beat.

Want a different match-up? Open the compare tool to add or swap models.

Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.