DeepSeek V4 Flash API pricing

Peak/off-peak pricing — schedule batch work off-peak to cut cost sharply. Available at listed at platform rate (no published official USD price).

In short: DeepSeek V4 Flash costs $0.22 per 1M input tokens and $0.66 per 1M output tokens. No official USD list price is published for this model, so it bills at platform rate — the reason to use it is access and consolidated billing rather than a discount. At 50M input and 10M output a month that is about $18. Rates verified 2026-09-16.
RouteInput / 1MOutput / 1M
Platform rate$0.22$0.66

USD, standard tier, verified 2026-09-16.

Check live rates on the platform →Free to sign up · $1 minimum top-up · No prepayment

Specifications

VendorDeepSeek
Context window
Max output
CategoryChina

What DeepSeek V4 Flash costs at your volume

Usage levelInput / output per monthMonthly cost
Light — prototyping, a few thousand calls5M / 1M$1.76
Working — one developer, daily use50M / 10M$17.60
Heavy — team or agent loops in production500M / 100M$176.00

Calculated at $0.22 input / $0.66 output per 1M tokens, assuming a 5:1 input-to-output ratio. Your ratio decides the real number — run your own figures through the cost calculator.

When DeepSeek V4 Flash is the right call

At $0.22 per million input tokens this sits in the low-cost band. It is a good fit for high-volume work — classification, extraction, summarisation at scale, autocomplete, and any pipeline where you call the model thousands of times a day.

It is usually the wrong call for tasks where a wrong answer is expensive to unwind; at this price the saving is small enough that escalating a hard case to a stronger model is usually worth it.

Similarly priced alternatives

Similarly pricedVendorIn / out per 1Mvs this
MiniMax M3MiniMax$0.15 / $0.6−$0.13
GPT-5.6 LunaOpenAI$0.1 / $0.6−$0.18
Gemini 3.7 FlashGoogle$0.375 / $1.875+$1.37

Blended comparison (input + output per 1M). Price is one axis — the models above are not interchangeable, they are the ones worth testing side by side at this budget.

Setting it up

Pointing an existing integration at this model is a base URL change, not a rewrite. Pick your tool:

FAQ

How much does DeepSeek V4 Flash cost per million tokens?

$0.22 per million input tokens and $0.66 per million output tokens. No official USD list price is published for this model, so it is billed at platform rate.

Is DeepSeek V4 Flash cheaper than buying direct?

There is no published official USD price to compare against for this model — it is billed at platform rate, and the reason to route it through a gateway is access and consolidated billing rather than a discount.

What does DeepSeek V4 Flash cost per month in practice?

At a 5:1 input-to-output ratio it is about $1.76 a month for light use (5M input tokens) and roughly $176.00 for heavy use (500M input). Agent workloads sit at the high end because every turn resends the context — see what actually drives Claude Code cost.

Which models cost about the same as DeepSeek V4 Flash?

The comparison table further up this page lists the three closest in blended price. If you are choosing on cost alone, start there and test the top two on your own prompts — price per token is only half the equation, the other half is how many tokens a model needs to finish the task.

How do I start using DeepSeek V4 Flash?

You need an API key, then point your client at the gateway base URL. It is a one-line change in most SDKs — see the SDK setup guide, or the Claude Code and Cursor guides for those tools.

Get DeepSeek V4 Flash access