Introducing Gemini 2.5 ProGoogleReleased June 27, 2025

Gemini 2.5 Pro

Google’s frontier reasoning and coding model with 2M token context, designed for complex analytical and mathematical workflows.

text-to-textProprietary API$1.25 / 1M tokContext: 2M (2,000,000 tokens)Arena ELO: 1355

Technical Specifications

Architecture Type
Frontier Multimodal MoE
Total Parameters
Undisclosed
Context Window
2M (2,000,000 tokens)
Max Output Tokens
16.384K (16,384 tokens)
Knowledge Cutoff
January 2025
Supported Modalities
text, code, vision, audio, video
License & Access
Google Cloud API Terms of Service

Benchmark Evaluations

Gpqa
64.2
Math500
88.5
Mmlu Pro
76.5
Humaneval
92.4
Swe Bench Verified
48.9
Chatbot Arena ELO
1355

Deep Architectural Overview

Gemini 2.5 Pro brings advanced mathematical proofs, full codebase refactoring, and multi-step reasoning to Google’s flagship tier. It integrates extended thinking capabilities while preserving the full 2-million-token input window.

Strengths & Considerations

Core Strengths
  • Frontier coding & SWE-bench performance
  • Unmatched 2M multimodal context window
  • Superior complex instructions following
Known Limitations
  • Higher latency than Flash tier models

Token & API Pricing

Input Tokens (1M)$1.25
Output Tokens (1M)$5.00
Cached Input (1M)$0.3125
Pricing is verified directly against Google's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Google and comparable reasoning engines.

Browse all models
Google

The latest evolution in the Gemini 3 family, delivering state-of-the-art software engineering (73.7% DeepSWE) and agentic enterprise knowledge workflows.

1.049M ctx$0.75 / 1M tok ($1.50 reg)
Google

Universal omnimodal generation model that accepts any combination of text, audio, image, and video to generate any combination of outputs.

1.049M ctx$0.60 / 1M tok (Universal Omni)