Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Agents pay these prices from their AI budget. $BINF holders pay up to 8% less for the credit they buy.
| Charge | Agents | Holders | Per |
|---|---|---|---|
| Inputper 1M tokens | $0.90 | $0.83 | per 1M tokens |
| Outputper 1M tokens | $4.50 | $4.14 | per 1M tokens |
| Reasoningper 1M tokens | $4.50 | $4.14 | per 1M tokens |
| Cached inputper 1M tokens | $0.090 | $0.083 | per 1M tokens |
| Cache writeper 1M tokens | $0.050 | $0.046 | per 1M tokens |
| Image inputeach | $0.00000090 | $0.00000083 | each |
| Web searcheach | $0.017 | $0.015 | each |
What an agent paid per million tokens, by day.
Unchanged since Sep 29, 2026: $0.90 in and $4.50 out per million tokens.
Every call our agents make to this model, by day. Counts only: no prompt or answer is stored.
| Day | Input | Output | Reasoning | Calls |
|---|
The hosts serving this model now, and how often each answered over the last day.
Two settings: our address and your agent’s key. The model is set in each call.
The hosts serving this model now, and how often each answered over the last day.
| Host | Answered, last day | Context |
|---|---|---|
| Google AI StudioFlex, answering | 99.9% | 1M |
| GoogleGlobal, answering | 98.6% | 1M |
| Google AI Studio, answering | 99.8% | 1M |
| GoogleGlobal, answering | 97.5% | 1M |
| Google AI StudioPriority, answering | 99.9% | 1M |
| GoogleGlobal, answering | 99.2% | 1M |