Gemini 2.5 Flash-Lite
Google's fastest and most budget-friendly model in the 2.5 family, built for scale.
Pricing
Input$0.10 text/image/video, $0.30 audio
Cached input$0.01 text/image/video, $0.03 audio
Output$0.40
Batch support and grounding are available.
Cost bandBudget
Free tierFree tier available with limits.
Limits
Context window1,048,576 checked Jul 19, 2026
Max output65,536
Knowledge cutoffJanuary 2025
DeploymentGemini Developer API and Google AI Studio
OpenAI-compatible APINo
ModalitiesText, Image, Audio, Video, PDF input
Where it fits
Best forMassive throughput, light reasoning, extraction, and multimodal utility pipelines.
Business fitA strong cost floor for teams that still want a mainstream managed model with 1M context.
Capabilities
ThinkingCachingBatch APICode executionFile searchFunction callingGroundingStructured outputs
Benchmark scores
Composite63 / 100
Coding60 / 100 (HumanEval (vendor))
Reasoning62 / 100 (MMLU-Pro (vendor))
Long context65 / 100 (RULER (vendor))
Vision66 / 100 (MMMU (vendor))
Instruction following64 / 100 (IFEval (vendor))
Editorial aggregate of published benchmark results as of Aug 1, 2025. These are not independent measurements run by this site.
Change history
No recorded changes for this model yet. Pricing and context are checked daily; changes appear here with dates and sources.
Data
Last checkedJul 19, 2026
Cross-checked againstopenrouter.ai/google/gemini-2.5-flash-lite
Verify with provider documentation before committing spend. Report an error and it gets fixed with a logged correction.