DeepSeek model
DeepSeek V4.1 Flash API cost calculator
DeepSeek V4.1 Flash costs $0.00 per 1M input tokens and $2.40 per 1M output tokens on the DeepSeek API, with a 1.0M-token context window. A typical request (2K input, 500 output tokens) costs $0.0012 — about $36.18/month at 1,000 requests per day. Use the calculator below to model your exact workload.
Find me a cheaper modelWhat real workloads cost on DeepSeek V4.1 Flash
| Workload | Tokens (in / out) | Per request | Per 1K requests |
|---|---|---|---|
| Chatbot message | 500 / 300 | $0.0007 | $0.72 |
| RAG query with context | 4,000 / 500 | $0.0012 | $1.21 |
| Document summarization | 20,000 / 1,000 | $0.0025 | $2.46 |
| Agent coding session | 100,000 / 5,000 | $0.01 | $12.30 |
DeepSeek V4.1 Flash vs other DeepSeek models
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| DeepSeek V4.1 Flash | $0.00 | $2.40 | 1.0M |
| DeepSeek V4 Flash Vision Exp | $0.22 | $0.65 | 1.0M |
| DeepSeek V4 Pro 0813 | $0.85 | $5 | 1.0M |
| DeepSeek V4 Flash 0731 | $0.02 | $1.28 | 1.0M |
| DeepSeek V4 Pro 0423 | $0.21 | $0.42 |
Frequently asked questions
How much does DeepSeek V4.1 Flash cost per 1M tokens?
According to current pricing data, DeepSeek V4.1 Flash costs $0.00 per 1M input tokens and $2.40 per 1M output tokens. Processing 1M tokens each way costs $2.40.
What does a typical API request to DeepSeek V4.1 Flash cost?
A typical request with 2,000 input tokens and 500 output tokens costs $0.0012. At 1,000 requests per day that is $1.21 daily, or about $36.18 per month.
What is DeepSeek V4.1 Flash's context window?
DeepSeek V4.1 Flash supports a 1.0M-token context window (1,048,576 tokens).
Does DeepSeek V4.1 Flash support prompt caching?
Yes. Cached input tokens cost $0.00 per 1M — a 0% discount versus the standard input rate, which matters for agents and chat apps that resend the same system prompt.
Prices may differ from DeepSeek's official pricing page for batch, tiered, or fine-tuned usage. Compare all current prices on the live model pricing dashboard, or run DeepSeek V4.1 Flash head-to-head against other models in the side-by-side comparison playground.
Pricing data via the OpenRouter models API.