Gemini 2.5 Flash-Lite

Google · Gemini 2.5 · budget · Stable model, latest update July 2025 · gemini-2.5-flash-lite

Gemini 2.5 family:Gemini 2.5 Pro·Gemini 2.5 Flash·Gemini 2.5 Flash-Lite

Google's fastest and most budget-friendly model in the 2.5 family, built for scale.

Pricing

Input$0.10 text/image/video, $0.30 audio
Cached input$0.01 text/image/video, $0.03 audio
Output$0.40

Batch support and grounding are available.

Cost bandBudget
Free tierFree tier available with limits.

Limits

Context window1,048,576 checked Jul 19, 2026
Max output65,536
Knowledge cutoffJanuary 2025
DeploymentGemini Developer API and Google AI Studio
OpenAI-compatible APINo
ModalitiesText, Image, Audio, Video, PDF input

Where it fits

Best forMassive throughput, light reasoning, extraction, and multimodal utility pipelines.
Business fitA strong cost floor for teams that still want a mainstream managed model with 1M context.

Capabilities

ThinkingCachingBatch APICode executionFile searchFunction callingGroundingStructured outputs

Benchmark scores

Composite63 / 100
Coding60 / 100 (HumanEval (vendor))
Reasoning62 / 100 (MMLU-Pro (vendor))
Long context65 / 100 (RULER (vendor))
Vision66 / 100 (MMMU (vendor))
Instruction following64 / 100 (IFEval (vendor))

Editorial aggregate of published benchmark results as of Aug 1, 2025. These are not independent measurements run by this site.

Change history

No recorded changes for this model yet. Pricing and context are checked daily; changes appear here with dates and sources.

Data

Last checkedJul 19, 2026
Cross-checked againstopenrouter.ai/google/gemini-2.5-flash-lite

Verify with provider documentation before committing spend. Report an error and it gets fixed with a logged correction.

Open in the comparison table JSON API