Introducing Gemini 3.1 ProGoogleReleased February 19, 2026

Gemini 3.1 Pro

Google’s most advanced model for complex coding, scientific discovery, and long-horizon multimodal reasoning with a 64K output token buffer.

text-to-textProprietary API$1.50 / 1M tok (64k output)Context: 1.049M (1,048,576 tokens)Arena ELO: 1435

Technical Specifications

Architecture Type
Frontier Multimodal Reasoning MoE
Total Parameters
Frontier Scale
Context Window
1.049M (1,048,576 tokens)
Max Output Tokens
65.536K (65,536 tokens)
Knowledge Cutoff
November 2025
Supported Modalities
text, code, vision, audio, video
License & Access
Google Cloud API Terms of Service

Benchmark Evaluations

Math500
94.7
Mmlu Pro
83.9
Humaneval
95.8
Swe Bench Verified
61.2
Chatbot Arena ELO
1435

Deep Architectural Overview

Gemini 3.1 Pro represents an architectural refinement on Gemini 3 Pro. It expands single-turn output capacity to 64k tokens, allowing complete multi-file codebases and comprehensive research syntheses to be generated without truncation.

Strengths & Considerations

Core Strengths
  • Massive 64k token single-turn generation capacity
  • Top-ranked SWE-bench score (61.2%)
  • Exceptional document & technical diagram comprehension
Known Limitations
  • Higher latency on long reasoning queries

Token & API Pricing

Input Tokens (1M)$1.50
Output Tokens (1M)$6.00
Cached Input (1M)$0.3750
Pricing is verified directly against Google's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Google and comparable reasoning engines.

Browse all models
Google

The latest evolution in the Gemini 3 family, delivering state-of-the-art software engineering (73.7% DeepSWE) and agentic enterprise knowledge workflows.

1.049M ctx$0.75 / 1M tok ($1.50 reg)
Google

Universal omnimodal generation model that accepts any combination of text, audio, image, and video to generate any combination of outputs.

1.049M ctx$0.60 / 1M tok (Universal Omni)