Gemini 3.8 Flash is our most intelligent workhorse model yet, delivering significant improvements from 3.7 Flash across software engineering, agentic tasks, and critical, multi-step reasoning in specialized domains. 3.8 Flash delivers substantial gains from 3.7 Flash, often approaching the performance of higher-cost frontier models.
At times, 3.8 Flash might use more tokens to maximize performance, especially at higher effort levels. For applications where compute efficiency is the primary constraint, you can use lower effort levels to minimize token overhead or use 3.7 Flash instead.
Try in Agent Studio Deploy example app Developer guide Pricing
| Model ID | gemini-3.8-flash |
|
|---|---|---|
| Modalities |
|
|
| Token limits | Context window | 1,048,576 |
| Maximum output tokens | 65,536 | |
| Capabilities |
|
|
| Tools |
|
|
| Consumption options |
|
|
| Technical specifications | Image |
|
| Text |
|
|
| Video |
|
|
| Audio |
|
|
| Parameter defaults |
|
|
| Supported regions |
|
|
|
||
|
||
|
||
| Versions |
|
|
| Security controls | Online prediction |
|
| Batch inference |
|
|
| Context caching |
|
|
| See Security controls for more information. | ||