Frequently asked questions
How much does the DeepSeek V4 Flash API cost?
DeepSeek V4 Flash costs $0.14 per million input tokens and $0.28 per million output tokens, with cached input at $0.0028 per million (98% off). Note: Cached-input price is DeepSeek’s cache-hit rate.
How much does a typical chat message cost with DeepSeek V4 Flash?
A short chat message (about 500 input and 250 output tokens) costs roughly $0.0001 with DeepSeek V4 Flash. At 1,000 such requests a day, that is about $4.20 per month.
Is DeepSeek V4 Flash cheap compared to similar models?
DeepSeek V4 Flash has the #3 cheapest input rate of the 26 models we track. On a mixed workload (10K input / 2K output), it costs $0.0020 per request, versus $0.0013 for GPT-5 nano and $0.0027 for GPT-4o mini.
Related
Compare all models at once ·Count tokens in your prompt · Run DeepSeek V4 Flash locally — VRAM requirements
Last updated 2026-08-03. Prices verified againstDeepSeek's official pricing page; see themethodology.