Gemini 3.8 Flash pricing in 2026

What one task costs on Gemini 3.8 Flash

Per-million-token prices are hard to feel, so here are typical token counts for common jobs priced with Gemini 3.8 Flash. Figures are USD at the rates above and exclude prompt caching and batch discounts.

TaskTokens (in / out)Cost
Code review of a 500-line pull request60,000 / 4,000$0.060
Debug one failing test40,000 / 6,000$0.052
Generate a landing page in one shot3,000 / 30,000$0.115
Draft a 1,200-word article2,000 / 2,500$0.011
Answer a question from a 20-page document25,000 / 800$0.022
Multi-step agent run (10 tool calls)250,000 / 25,000$0.281
All of the above, once$0.541

Cheaper at the same or better capability: GLM-5.3 Flash runs the same set of tasks for $0.091 instead of $0.541.

Gemini 3.8 Flash by Google costs $0.75 per 1M input tokens and $3.75 per 1M output tokens — about 333.3k tokens per $1. Prices last verified 2026-09-19.

The verdict

Google's most intelligent Flash yet: an index of 41 at $0.75 / $3.75 with a 1M-token context and roughly 277 tokens per second. That combination — near-Pro quality, Flash pricing, top-tier speed — is why it currently posts one of the best intelligence-per-dollar scores on the site. The promotional price runs through 31 December 2026, after which it doubles to $1.50 / $7.50.

A concrete example: sending 100,000 input tokens and getting back 20,000 output tokens with Gemini 3.8 Flash costs about $0.15 at current prices.

Who is Gemini 3.8 Flash for?

Gemini 3.8 Flash is Google's most intelligent Flash model, with significant gains over 3.7 Flash across software engineering, agentic tasks and multi-step reasoning. At 41 on the Artificial Analysis Intelligence Index it sits above every other sub-$1 model we track, while running at around 277 tokens per second — among the fastest models on the market.

The catch is verbosity and iteration: it generates a lot of tokens for its answers, and it is stronger when you can take a couple of shots than when you need one perfect pass. Both are cheap to absorb at $0.75 / $3.75, which is a promotional rate through 31 December 2026 — plan for $1.50 / $7.50 after that. For long-context work, high-volume agents and anything latency-sensitive, it is currently the best value in the Google line.

Strengths & weaknesses

  • Strength: Near-Pro quality at Flash price and speed
  • Best for: Speed, Budget, Iteration, Long-context
  • Weakness: Very verbose (170M tokens to run the index) and needs multiple shots on one-shot coding tasks.

Gemini 3.8 Flash key facts

ProviderGemini pricing
Input price$0.75 / 1M tokens
Output price$3.75 / 1M tokens
Tokens / $1333.3k
Intelligence / $1136,667
Context window1.0M tokens
Tierbalanced
SpeedFastest (277+ tokens/sec)
Intelligence index41 / 100
Released2026-09-02
Last verified2026-09-19

What Gemini 3.8 Flash delivers for your budget

For a coding workload, here is roughly what each monthly budget buys:

BudgetTokens / monthMonths of workload
$2013.9M21.3x
$5034.7M53.3x
$10069.3M106.7x

Other Gemini models

How Gemini 3.8 Flash compares to its closest rivals

Rather than repeating the full catalog on every page, here are the three current models nearest to Gemini 3.8 Flash on the intelligence index — and why you might still pick one instead:

  • GPT-5.6 Terra (ChatGPT) — $2 input / $12 output per 1M tokens. Why you might pick it: Best ChatGPT value for most tasks.
  • DeepSeek V4.1 Flash (DeepSeek) — $0.15 input / $0.60 output per 1M tokens. Why you might pick it: Best intelligence-per-dollar on the site.
  • Qwen3.8 2.4T A95B (Qwen) — $2 input / $6 output per 1M tokens. Why you might pick it: Frontier open-weight design.

Gemini 3.8 Flash — frequently asked questions

How much does Gemini 3.8 Flash cost?

Gemini 3.8 Flash by Google costs $0.75 per 1M input tokens and $3.75 per 1M output tokens.

How many tokens per dollar does Gemini 3.8 Flash give you?

On a typical 1:3 input-to-output mix, Gemini 3.8 Flash delivers around 333.3k tokens per dollar — an intelligence-per-dollar score of 136667.

What is Gemini 3.8 Flash's context window?

Gemini 3.8 Flash supports a context window of 1.0M tokens.

Is Gemini 3.8 Flash worth it in 2026?

Gemini 3.8 Flash is Gemini's balanced tier model. Use the AI Intelligence Calculator to see how its value compares to every other model at your budget.

Advertisement