google/gemini-3.1-flash-liteGemini 3.1 Flash-Lite is Google’s proprietary AI model, released Mar 3, 2026. It costs $0.25/M input and $1.50/M output tokens, and handles a 1M-token context window.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.
Scores on standardized evaluations. Higher is better — and rank shows where Gemini 3.1 Flash-Lite lands among all models tracked on Model Beat.
Graduate-level scientific reasoning
Frontier of human expert knowledge
Novel ML problem-solving
Scientific research coding
Human-rated web development
Multi-step agentic tasks
Tool-agent-user reliability
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.
Gemini 3.1 Flash-Lite is an AI model developed by Google, part of the Gemini family, released Mar 3, 2026.
Gemini 3.1 Flash-Lite's list price from Google is $0.25 per million input tokens and $1.50 per million output tokens.
Gemini 3.1 Flash-Lite supports a context window of up to 1M tokens.
On standardized evaluations tracked by Epoch AI, Gemini 3.1 Flash-Lite scores 52.2% on WeirdML, 41.9% on SciCode, 82.2% on GPQA Diamond.
Gemini 3.1 Flash-Lite is a proprietary model (API access), available through its provider's API rather than as a downloadable open-weight model.
Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Jul 30, 2026.