Home/DeepSeek API/DeepSeek V4.1 Flash
DeepSeek V4.1 Flash Pricing Calculator
DeepSeek's main API model (deepseek-flash), with thinking and non-thinking modes. One typical prompt costs $0.000750 (0.075¢) on DeepSeek V4.1 Flash. Paste your own prompt below to compare it with other models.
Peak-hour price shown. DeepSeek charges 50% less off-peak — every hour except 01:00–04:00 and 06:00–10:00 UTC on weekdays (about 79% of the week). Off-peak: $0.15 input / $0.60 output per 1M tokens.
Source: DeepSeek official pricing · verified October 2026.
How many tokens you expect the model to generate per response
Showing a typical prompt (500 input + 500 output tokens) across 26 models. Paste your own prompt above for an exact count.
| Model | Per Prompt | Prompts / $1 | Monthly |
|---|---|---|---|
| DeepSeek | $0.0007500.075¢ | 1,333 | $2.25 |
GPT-6 LunaCheapest OpenAI | $0.0003000.03¢ | 3,333 | $0.9000 |
| Mistral | $0.0003750.037¢ | 2,666 | $1.13 |
| OpenAI (via Together) | $0.0003750.037¢ | 2,666 | $1.13 |
| OpenAI | $0.0007000.07¢ | 1,428 | $2.10 |
| Meta (via Together) | $0.0009250.092¢ | 1,081 | $2.77 |
| Mistral | $0.00100.1¢ | 1,000 | $3.00 |
| $0.00140.14¢ | 714 | $4.20 | |
| xAI | $0.00190.19¢ | 533 | $5.63 |
| $0.00220.22¢ | 444 | $6.75 | |
| DeepSeek | $0.00260.26¢ | 378 | $7.92 |
| Meta | $0.00270.27¢ | 363 | $8.25 |
| Anthropic | $0.00300.3¢ | 333 | $9.00 |
| xAI | $0.00400.4¢ | 250 | $12.00 |
| Alibaba (Qwen) | $0.00400.4¢ | 250 | $12.00 |
| Mistral | $0.00450.45¢ | 222 | $13.50 |
| OpenAI | $0.00600.6¢ | 166 | $18.00 |
| OpenAI | $0.00600.6¢ | 166 | $18.00 |
| Anthropic | $0.00600.6¢ | 166 | $18.00 |
| OpenAI | $0.00700.7¢ | 142 | $21.00 |
| $0.00700.7¢ | 142 | $21.00 | |
| Moonshot AI | $0.00900.9¢ | 111 | $27.00 |
| OpenAI | $0.0120 | 83 | $36.00 |
| Anthropic | $0.0120 | 83 | $36.00 |
| OpenAI | $0.0300 | 33 | $90.00 |
| Anthropic | $0.0300 | 33 | $90.00 |
Prices from official provider pricing pages, verified October 2026.
How Much Does DeepSeek V4.1 Flash Cost?
| Prompts | Cost (500 in + 500 out tokens each) |
|---|---|
| 1 prompt | $0.000750 |
| 100 prompts | $0.0750 |
| 1,000 prompts | $0.7500 |
| 10,000 prompts | $7.50 |
| 100,000 prompts | $75.00 |
Embed this calculator on your site.
| Input price | $0.30 / 1M |
|---|---|
| Output price | $1.20 / 1M |
| Cached input | $0.006 / 1M |
| Typical prompt | $0.000750 |
| Batch discount | None |
| Context window | 1M tokens |
| Max output | 384,000 tokens |
| Knowledge cutoff | Not published |
| Input types | Text, Images |
| Reasoning mode | Optional |
| Open weights | Yes |
| Released | September 10, 2026 |
| Status | Current |
Comparisons with DeepSeek V4.1 Flash
DeepSeek V4.1 Flash — Frequently Asked Questions
A typical prompt of 500 input tokens with a 500-token answer costs $0.000750 (0.075¢) on DeepSeek V4.1 Flash. That means $1 buys about 1,333 prompts. Longer prompts and answers cost proportionally more.
With 500 input and 500 output tokens per request, DeepSeek V4.1 Flash costs $0.7500 per 1,000 requests. Input tokens are $0.30 per million and output tokens $1.20 per million.
DeepSeek V4.1 Flash is more expensive than GPT-5.6 Luna — about 1.1x more for the same request. DeepSeek V4.1 Flash costs $0.30/$1.20 per million tokens (input/output) vs GPT-5.6 Luna's $0.20/$1.20.
DeepSeek V4.1 Flash accepts text and images, writes up to 384,000 tokens per response, has an optional reasoning (thinking) mode and has open weights you can self-host. Its context window is 1M tokens.
DeepSeek V4.1 Flash supports 1M tokens (1,000,000) — the maximum combined length of your prompt and the model's answer in one request.
At 10,000 requests per day with 500 input + 500 output tokens each, DeepSeek V4.1 Flash costs about $225.00/month. Use the calculator above with your real prompt and volume.