google/gemini-2.5-flash-liteGemini 2.5 Flash-Lite is Google DeepMind’s proprietary AI model, released Sep 25, 2025. It costs $0.10/M input and $0.40/M output tokens, and handles a 1M-token context window.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.
Scores on standardized evaluations. Higher is better — and rank shows where Gemini 2.5 Flash-Lite lands among all models tracked on Model Beat.
Graduate-level scientific reasoning
Frontier of human expert knowledge
Broad expert knowledge
Novel ML problem-solving
Contamination-free coding
Scientific research coding
Olympiad-qualifier math
Tool-agent-user reliability
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.
Gemini 2.5 Flash-Lite is an AI model developed by Google DeepMind, part of the Gemini family, released Sep 25, 2025.
Gemini 2.5 Flash-Lite's list price from Google DeepMind is $0.10 per million input tokens and $0.40 per million output tokens.
Gemini 2.5 Flash-Lite supports a context window of up to 1M tokens.
On standardized evaluations tracked by Epoch AI, Gemini 2.5 Flash-Lite scores 59.3% on LiveCodeBench, 75.9% on MMLU-Pro, 6.4% on Humanity's Last Exam.
Gemini 2.5 Flash-Lite is a proprietary model (API access), available through its provider's API rather than as a downloadable open-weight model.
Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Jul 30, 2026.