Gemini 3.8 Flash pricing in 2026
What one task costs on Gemini 3.8 Flash
Per-million-token prices are hard to feel, so here are typical token counts for common jobs priced with Gemini 3.8 Flash. Figures are USD at the rates above and exclude prompt caching and batch discounts.
| Task | Tokens (in / out) | Cost |
|---|---|---|
| Code review of a 500-line pull request | 60,000 / 4,000 | $0.060 |
| Debug one failing test | 40,000 / 6,000 | $0.052 |
| Generate a landing page in one shot | 3,000 / 30,000 | $0.115 |
| Draft a 1,200-word article | 2,000 / 2,500 | $0.011 |
| Answer a question from a 20-page document | 25,000 / 800 | $0.022 |
| Multi-step agent run (10 tool calls) | 250,000 / 25,000 | $0.281 |
| All of the above, once | $0.541 |
Cheaper at the same or better capability: GLM-5.3 Flash runs the same set of tasks for $0.091 instead of $0.541.
Gemini 3.8 Flash by Google costs $0.75 per 1M input tokens and $3.75 per 1M output tokens — about 333.3k tokens per $1. Prices last verified 2026-09-19.
The verdict
Google's most intelligent Flash yet: an index of 41 at $0.75 / $3.75 with a 1M-token context and roughly 277 tokens per second. That combination — near-Pro quality, Flash pricing, top-tier speed — is why it currently posts one of the best intelligence-per-dollar scores on the site. The promotional price runs through 31 December 2026, after which it doubles to $1.50 / $7.50.
A concrete example: sending 100,000 input tokens and getting back 20,000 output tokens with Gemini 3.8 Flash costs about $0.15 at current prices.
Who is Gemini 3.8 Flash for?
Gemini 3.8 Flash is Google's most intelligent Flash model, with significant gains over 3.7 Flash across software engineering, agentic tasks and multi-step reasoning. At 41 on the Artificial Analysis Intelligence Index it sits above every other sub-$1 model we track, while running at around 277 tokens per second — among the fastest models on the market.
The catch is verbosity and iteration: it generates a lot of tokens for its answers, and it is stronger when you can take a couple of shots than when you need one perfect pass. Both are cheap to absorb at $0.75 / $3.75, which is a promotional rate through 31 December 2026 — plan for $1.50 / $7.50 after that. For long-context work, high-volume agents and anything latency-sensitive, it is currently the best value in the Google line.
Strengths & weaknesses
- Strength: Near-Pro quality at Flash price and speed
- Best for: Speed, Budget, Iteration, Long-context
- Weakness: Very verbose (170M tokens to run the index) and needs multiple shots on one-shot coding tasks.
Gemini 3.8 Flash key facts
| Provider | Gemini pricing |
|---|---|
| Input price | $0.75 / 1M tokens |
| Output price | $3.75 / 1M tokens |
| Tokens / $1 | 333.3k |
| Intelligence / $1 | 136,667 |
| Context window | 1.0M tokens |
| Tier | balanced |
| Speed | Fastest (277+ tokens/sec) |
| Intelligence index | 41 / 100 |
| Released | 2026-09-02 |
| Last verified | 2026-09-19 |
What Gemini 3.8 Flash delivers for your budget
For a coding workload, here is roughly what each monthly budget buys:
| Budget | Tokens / month | Months of workload |
|---|---|---|
| $20 | 13.9M | 21.3x |
| $50 | 34.7M | 53.3x |
| $100 | 69.3M | 106.7x |
Other Gemini models
How Gemini 3.8 Flash compares to its closest rivals
Rather than repeating the full catalog on every page, here are the three current models nearest to Gemini 3.8 Flash on the intelligence index — and why you might still pick one instead:
- GPT-5.6 Terra (ChatGPT) — $2 input / $12 output per 1M tokens. Why you might pick it: Best ChatGPT value for most tasks.
- DeepSeek V4.1 Flash (DeepSeek) — $0.15 input / $0.60 output per 1M tokens. Why you might pick it: Best intelligence-per-dollar on the site.
- Qwen3.8 2.4T A95B (Qwen) — $2 input / $6 output per 1M tokens. Why you might pick it: Frontier open-weight design.
Gemini 3.8 Flash — frequently asked questions
How much does Gemini 3.8 Flash cost?
Gemini 3.8 Flash by Google costs $0.75 per 1M input tokens and $3.75 per 1M output tokens.
How many tokens per dollar does Gemini 3.8 Flash give you?
On a typical 1:3 input-to-output mix, Gemini 3.8 Flash delivers around 333.3k tokens per dollar — an intelligence-per-dollar score of 136667.
What is Gemini 3.8 Flash's context window?
Gemini 3.8 Flash supports a context window of 1.0M tokens.
Is Gemini 3.8 Flash worth it in 2026?
Gemini 3.8 Flash is Gemini's balanced tier model. Use the AI Intelligence Calculator to see how its value compares to every other model at your budget.