google/gemini-3.1-flash-liteGemini 3.1 Flash-Lite is Google’s proprietary AI model, released Mar 3, 2026. It costs $0.28/M input and $1.65/M output tokens, and handles a 1M-token context window.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.
Scores on standardized evaluations. Higher is better — and rank shows where Gemini 3.1 Flash-Lite lands among all models tracked on Model Beat.
Graduate-level scientific reasoning
Frontier of human expert knowledge
Novel ML problem-solving
Scientific research coding
Human-rated web development
Olympiad-qualifier math
Multi-step agentic tasks
Tool-agent-user reliability
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.
Gemini 3.1 Flash-Lite is an AI model developed by Google, part of the Gemini family, released Mar 3, 2026.
Gemini 3.1 Flash-Lite's list price from Google is $0.28 per million input tokens and $1.65 per million output tokens. The cheapest credible third-party provider on OpenRouter (Google) serves it at $0.25/$1.50 per 1M.
Gemini 3.1 Flash-Lite supports a context window of up to 1M tokens.
On standardized evaluations tracked by Epoch AI, Gemini 3.1 Flash-Lite scores 43.4% on SciCode, 52.2% on WeirdML, 81.8% on GPQA Diamond.
Gemini 3.1 Flash-Lite is a proprietary model (API access), available through its provider's API rather than as a downloadable open-weight model.
Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Sep 17, 2026.