The latest voice foundation model from Google DeepMind, integrating extended chain-of-thought reasoning directly into live spoken dialogues.
Gemini 3.6 Flash
Iterative performance leap in the Flash family featuring major upgrades in code refactoring and autonomous agent tool loops.
Technical Specifications
Benchmark Evaluations
Deep Architectural Overview
Gemini 3.6 Flash bridges the gap between speed and frontier-grade capability, bringing 94.6% HumanEval performance and 64k token outputs to the high-efficiency Flash tier.
Strengths & Considerations
- 94.6% HumanEval code generation
- 64k token single-response output
- Excellent tool calling consistency
- Slightly higher cost than 3.5 Flash
Token & API Pricing
Similar & Alternative Models
Explore other frontier models from Google and comparable reasoning engines.
The latest evolution in the Gemini 3 family, delivering state-of-the-art software engineering (73.7% DeepSWE) and agentic enterprise knowledge workflows.
Universal omnimodal generation model that accepts any combination of text, audio, image, and video to generate any combination of outputs.
Universal multilingual speech engine supporting live simultaneous translation and speaker diarization across 100+ languages.