Home/DeepSeek API/DeepSeek V4.1 Flash

DeepSeek V4.1 Flash Pricing Calculator

DeepSeek's main API model (deepseek-flash), with thinking and non-thinking modes. One typical prompt costs $0.000750 (0.075¢) on DeepSeek V4.1 Flash. Paste your own prompt below to compare it with other models.

Input Price
$0.30
per 1M tokens
Output Price
$1.20
per 1M tokens
Cached Input
$0.006
per 1M tokens
Context Window
1M
DeepSeek · September 10, 2026

Peak-hour price shown. DeepSeek charges 50% less off-peak — every hour except 01:00–04:00 and 06:00–10:00 UTC on weekdays (about 79% of the week). Off-peak: $0.15 input / $0.60 output per 1M tokens.

Source: DeepSeek official pricing · verified October 2026.

0input tokens
Estimated (~4 chars/token)

How many tokens you expect the model to generate per response

Showing a typical prompt (500 input + 500 output tokens) across 26 models. Paste your own prompt above for an exact count.

ModelPer PromptPrompts / $1Monthly
DeepSeek$0.0007500.075¢1,333$2.25
OpenAI$0.0003000.03¢3,333$0.9000
Mistral$0.0003750.037¢2,666$1.13
OpenAI (via Together)$0.0003750.037¢2,666$1.13
OpenAI$0.0007000.07¢1,428$2.10
Meta (via Together)$0.0009250.092¢1,081$2.77
Mistral$0.00100.1¢1,000$3.00
Google$0.00140.14¢714$4.20
xAI$0.00190.19¢533$5.63
Google$0.00220.22¢444$6.75
DeepSeek$0.00260.26¢378$7.92
Meta$0.00270.27¢363$8.25
Anthropic$0.00300.3¢333$9.00
xAI$0.00400.4¢250$12.00
Alibaba (Qwen)$0.00400.4¢250$12.00
Mistral$0.00450.45¢222$13.50
OpenAI$0.00600.6¢166$18.00
OpenAI$0.00600.6¢166$18.00
Anthropic$0.00600.6¢166$18.00
OpenAI$0.00700.7¢142$21.00
Google$0.00700.7¢142$21.00
Moonshot AI$0.00900.9¢111$27.00
OpenAI$0.012083$36.00
Anthropic$0.012083$36.00
OpenAI$0.030033$90.00
Anthropic$0.030033$90.00

Prices from official provider pricing pages, verified October 2026.

DeepSeek V4.1 Flash Specs & Pricing
Input price$0.30 / 1M
Output price$1.20 / 1M
Cached input$0.006 / 1M
Typical prompt$0.000750
Batch discountNone
Context window1M tokens
Max output384,000 tokens
Knowledge cutoffNot published
Input typesText, Images
Reasoning modeOptional
Open weightsYes
ReleasedSeptember 10, 2026
StatusCurrent

DeepSeek V4.1 Flash — Frequently Asked Questions

A typical prompt of 500 input tokens with a 500-token answer costs $0.000750 (0.075¢) on DeepSeek V4.1 Flash. That means $1 buys about 1,333 prompts. Longer prompts and answers cost proportionally more.

With 500 input and 500 output tokens per request, DeepSeek V4.1 Flash costs $0.7500 per 1,000 requests. Input tokens are $0.30 per million and output tokens $1.20 per million.

DeepSeek V4.1 Flash is more expensive than GPT-5.6 Luna — about 1.1x more for the same request. DeepSeek V4.1 Flash costs $0.30/$1.20 per million tokens (input/output) vs GPT-5.6 Luna's $0.20/$1.20.

DeepSeek V4.1 Flash accepts text and images, writes up to 384,000 tokens per response, has an optional reasoning (thinking) mode and has open weights you can self-host. Its context window is 1M tokens.

DeepSeek V4.1 Flash supports 1M tokens (1,000,000) — the maximum combined length of your prompt and the model's answer in one request.

At 10,000 requests per day with 500 input + 500 output tokens each, DeepSeek V4.1 Flash costs about $225.00/month. Use the calculator above with your real prompt and volume.