DeepSeek V4 Flash API Cost Calculator

DeepSeek's DeepSeek V4 Flash: $0.14/M input · $0.28/M output · $0.0028/M cached input. Price your own workload below.

Cached-input price is DeepSeek’s cache-hit rate.

Presets:

What DeepSeek V4 Flash costs in practice

WorkloadInputOutputCost
Short chat message500250$0.0001
10-page document + summary7,000500$0.0011
Agent / coding session50,00010,000$0.0098
1M tokens in + 1M out1,000,0001,000,000$0.420

DeepSeek V4 Flash vs nearest rivals

Closest-priced alternatives on a mixed 10K-input / 2K-output request:

ModelInput $/MTokOutput $/MTokMixed request
DeepSeek V4 Flash$0.14$0.28$0.0020
GPT-5 nano$0.05$0.4$0.0013
Gemini 2.5 Flash-Lite$0.1$0.4$0.0018
GPT-4o mini$0.15$0.6$0.0027

Frequently asked questions

How much does the DeepSeek V4 Flash API cost?

DeepSeek V4 Flash costs $0.14 per million input tokens and $0.28 per million output tokens, with cached input at $0.0028 per million (98% off). Note: Cached-input price is DeepSeek’s cache-hit rate.

How much does a typical chat message cost with DeepSeek V4 Flash?

A short chat message (about 500 input and 250 output tokens) costs roughly $0.0001 with DeepSeek V4 Flash. At 1,000 such requests a day, that is about $4.20 per month.

Is DeepSeek V4 Flash cheap compared to similar models?

DeepSeek V4 Flash has the #3 cheapest input rate of the 26 models we track. On a mixed workload (10K input / 2K output), it costs $0.0020 per request, versus $0.0013 for GPT-5 nano and $0.0027 for GPT-4o mini.

Related

Compare all models at once ·Count tokens in your prompt · Run DeepSeek V4 Flash locally — VRAM requirements

Last updated 2026-08-03. Prices verified againstDeepSeek's official pricing page; see themethodology.