Gemini 2.5 Pro

Google · Gemini 2.5 · frontier · Stable model, latest update June 2025 · gemini-2.5-pro

Gemini 2.5 family:Gemini 2.5 Pro·Gemini 2.5 Flash·Gemini 2.5 Flash-Lite

Google's advanced long-context thinking model for code, STEM, large datasets, and document-heavy work.

Pricing

Input$1.25 ≤200k, $2.50 >200k
Cached input$0.125 ≤200k, $0.25 >200k
Output$10 ≤200k, $15 >200k

Thinking tokens are included in output pricing; storage fees apply for context cache.

Cost bandPremium
Free tierFree tier in AI Studio and Gemini Developer API with limits; paid tier for production.

Limits

Context window1,048,576 checked Jul 19, 2026
Max output65,536
Knowledge cutoffJanuary 2025
DeploymentGemini Developer API and Google AI Studio
OpenAI-compatible APINo
ModalitiesText, Image, Audio, Video, PDF input

Where it fits

Best forLarge-context reasoning, codebase analysis, long documents, multimodal enterprise research.
Business fitOften the strongest Google pick when quality and long context beat cost concerns.

Capabilities

ThinkingCachingBatch APICode executionFile searchFunction callingSearch groundingStructured outputs

Benchmark scores

Composite89 / 100
Coding88 / 100 (LiveCodeBench)
Reasoning93 / 100 (GPQA Diamond)
Long context94 / 100 (RULER 1M)
Vision91 / 100 (MMMU)
Instruction following88 / 100 (IFEval)

Editorial aggregate of published benchmark results as of Sep 1, 2025. These are not independent measurements run by this site.

Change history

No recorded changes for this model yet. Pricing and context are checked daily; changes appear here with dates and sources.

Data

Last checkedJul 19, 2026
Cross-checked againstopenrouter.ai/google/gemini-2.5-pro

Verify with provider documentation before committing spend. Report an error and it gets fixed with a logged correction.

Open in the comparison table JSON API