DeepSeek V4 Flash pricing (retired)

Retired model. DeepSeek has superseded DeepSeek V4 Flash — the prices below are the last ones we verified and are kept for reference. Its replacement is DeepSeek V4.1 Flash ($0.15 / $0.60 per 1M in / out).

What one task cost on DeepSeek V4 Flash

Per-million-token prices are hard to feel, so here are typical token counts for common jobs priced with DeepSeek V4 Flash. Figures are USD at the rates above and exclude prompt caching and batch discounts.

TaskTokens (in / out)Cost
Code review of a 500-line pull request60,000 / 4,000$0.016
Debug one failing test40,000 / 6,000$0.013
Generate a landing page in one shot3,000 / 30,000$0.020
Draft a 1,200-word article2,000 / 2,500$0.0021
Answer a question from a 20-page document25,000 / 800$0.0060
Multi-step agent run (10 tool calls)250,000 / 25,000$0.072
All of the above, once$0.129

DeepSeek V4 Flash by DeepSeek costs $0.22 per 1M input tokens and $0.66 per 1M output tokens — about 1.8M tokens per $1. Prices last verified 2026-09-19 (historical).

The verdict

RETIRED — DeepSeek's documentation now states that the v4-flash name is legacy and that requests are served by DeepSeek V4.1 Flash ($0.15 / $0.60 off-peak). The entry is kept for historical reference; use [DeepSeek V4.1 Flash](/models/deepseek-v4-1-flash/) instead.

A concrete example: sending 100,000 input tokens and getting back 20,000 output tokens with DeepSeek V4 Flash costs about $0.04 at current prices.

Who is DeepSeek V4 Flash for?

DeepSeek V4 Flash was one of the cheaper APIs we tracked — about $0.22 / $0.66 per 1M tokens off-peak — while still carrying a then-near-frontier index of 35. It was the pick for bulk generation, batch jobs and cost-dominant workloads where token volume is everything.

In September 2026 DeepSeek retired the name: the v4-flash endpoint still answers, but requests are now served by DeepSeek V4.1 Flash at the same Flash pricing. For new work, use V4.1 Flash — it is cheaper ($0.15 / $0.60 off-peak) and scores higher on the current index.

Strengths & weaknesses

  • Strength: Historic: rock-bottom price at launch
  • Best for: Budget, Bulk, Local
  • Weakness: Retired and superseded by DeepSeek V4.1 Flash.

DeepSeek V4 Flash key facts

ProviderDeepSeek pricing
Input price$0.22 / 1M tokens
Output price$0.66 / 1M tokens
Tokens / $11.8M
Intelligence / $1636,364
Context window1.3M tokens
Tierbalanced
SpeedSlow
Intelligence index35 / 100
Released2026-07-31
Last verified2026-09-19

What DeepSeek V4 Flash delivers for your budget

For a coding workload, here is roughly what each monthly budget buys:

BudgetTokens / monthMonths of workload
$2062.2M95.7x
$50155.5M239.2x
$100311.0M478.5x

Other DeepSeek models

How DeepSeek V4 Flash compares to its closest rivals

Rather than repeating the full catalog on every page, here are the three current models nearest to DeepSeek V4 Flash on the intelligence index — and why you might still pick one instead:

  • DeepSeek V4 Pro (DeepSeek) — $0.66 input / $1.98 output per 1M tokens. Why you might pick it: Unbeatable frontier value.
  • GPT-5.6 Luna (ChatGPT) — $0.20 input / $1.20 output per 1M tokens. Why you might pick it: ChatGPT quality at a budget price.
  • Claude Sonnet 5 (Claude) — $2 input / $10 output per 1M tokens. Why you might pick it: Strong coding at mid-tier price.

DeepSeek V4 Flash — frequently asked questions

How much does DeepSeek V4 Flash cost?

DeepSeek V4 Flash by DeepSeek costs $0.22 per 1M input tokens and $0.66 per 1M output tokens.

How many tokens per dollar does DeepSeek V4 Flash give you?

On a typical 1:3 input-to-output mix, DeepSeek V4 Flash delivers around 1.8M tokens per dollar — an intelligence-per-dollar score of 636364.

What is DeepSeek V4 Flash's context window?

DeepSeek V4 Flash supports a context window of 1.3M tokens.

Is DeepSeek V4 Flash worth it in 2026?

DeepSeek V4 Flash is DeepSeek's balanced tier model. Use the AI Intelligence Calculator to see how its value compares to every other model at your budget.

Advertisement