The latest voice foundation model from Google DeepMind, integrating extended chain-of-thought reasoning directly into live spoken dialogues.
Gemini 3.8 Flash
The latest evolution in the Gemini 3 family, delivering state-of-the-art software engineering (73.7% DeepSWE) and agentic enterprise knowledge workflows.
Technical Specifications
Benchmark Evaluations
Deep Architectural Overview
Gemini 3.8 Flash delivers substantial performance advancements across software engineering and agentic knowledge workflows, outperforming Claude Sonnet 5 and GPT-5.6 on DeepSWE v1.1 and agentic terminal coding.
Strengths & Considerations
- 73.7% on DeepSWE v1.1 long-horizon software engineering
- 89.4% on Terminal-bench 2.1 agentic coding
- Customizable effort levels balancing quality, latency, and cost
- Complex legal all-pass rates remain challenging across industry
Token & API Pricing
Similar & Alternative Models
Explore other frontier models from Google and comparable reasoning engines.
Universal omnimodal generation model that accepts any combination of text, audio, image, and video to generate any combination of outputs.
Universal multilingual speech engine supporting live simultaneous translation and speaker diarization across 100+ languages.
High-performance multimodal foundation model featuring algorithmic reasoning enhancements and agentic long-form video understanding.