Primary use stated or inferred cautiously from official documentation.
Gemini 3.8 Flash
Flash model for long-running software engineering, autonomous agents, and complex business flows.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Limits and integration
| API ID | gemini-3.8-flash |
|---|---|
| Release | September 2, 2026 |
| Knowledge cutoff | Not published in the source consulted |
| Entrance | Text · Image · Video · Audio · PDF |
| Exit | Text |
| Context window | 1,048,576 entry tokens |
| maximum output | 65,536 output tokens |
| Reasoning | low, medium, and high levels |
| Published tools | Features · Web search · Google Maps · Code execution · Computer use (preview) · File search · URL context |
| Structured Outings | Compatible |
| Batch processing | Compatible · 50% discount |
| Prompt cache | Compatible |
| Fine-tuning | Not published in the source consulted |
| Verified platforms | Gemini API · Google AI Studio |
Low-cost agents, long-running software, and broad multimodal analysis.
The price is introductory, and using grounding may incur per-query charges.
What it can do
Freelance agents
Planning and orchestrating tools into multi-step objectives.
Code
Aimed at extensive software engineering and extensive refactorings.
Multimodality
Supports text, image, video, audio and PDF as input.
Adjustable reasoning
It offers low, medium and high levels to adjust the effort.
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | 0.75 USD / 1M tokens | Standard API rate |
| Cached input | 0.075 USD / 1 M tokens | Reading reused prefixes |
| Cache write or storage | 0.50 USD / 1 M tokens per hour of storage | The condition varies by provider |
| Standard output | 3.75 USD / 1M tokens | May include reasoning tokens |
| Batch input | 0.375 USD / 1 M tokens | Asynchronous processing |
| Batch output | 1.875 USD / 1 M tokens | Asynchronous processing |
Introductory pricing until December 31, 2026; from 2027, standard rates double.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
Google describes it as its smartest Flash model for code and agents. Prices shown are introductory until December 31, 2026.
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
Gemini 3.8 Flash
General availability and stable version.
Gemini 3.7 Flash
Previous version for code and agents.
Gemini 3.6 Flash
First stable version of the series 3.6.