Qwen2.5 VL 72B Instruct vs DeepSeek Flash Latest: pricing comparison
DeepSeek Flash Latest is the cheaper option at $0.00/$2.40 per 1M input/output tokens — Qwen2.5 VL 72B Instruct ($0.80/$1) costs about 267x more per input token. Full spec-by-spec breakdown below.
| Qwen2.5 VL 72B Instruct | DeepSeek Flash Latest | |
|---|---|---|
| Input /1M tokens | $0.80 | $0.00 |
| Output /1M tokens | $1 | $2.40 |
| Cache read /1M | $0.40 | $0.00 |
| Context window | 128K | 1.0M |
| Provider | Qwen | ~deepseek |
| Vision input | Yes | Yes |
| Released | Feb 2025 | Sep 2026 |
Real workload costs: Qwen2.5 VL 72B Instruct vs DeepSeek Flash Latest
Cost per single request at common token profiles — the cheapest model for each workload is highlighted.
| Workload | Tokens (in / out) | Qwen2.5 VL 72B Instruct | DeepSeek Flash Latest |
|---|---|---|---|
| Chatbot message | 500 / 300 | $0.0007 | $0.0007 |
| RAG query with context | 4,000 / 500 | $0.0037 | $0.0012 |
| Document summarization | 20,000 / 1,000 | $0.02 | $0.0025 |
| Agent coding session | 100,000 / 5,000 | $0.09 | $0.01 |
Frequently asked questions
Which is cheaper: Qwen2.5 VL 72B Instruct vs DeepSeek Flash Latest?
DeepSeek Flash Latest is the cheapest on input tokens at $0.00 per 1M, and Qwen2.5 VL 72B Instruct is cheapest on output at $1 per 1M. Qwen2.5 VL 72B Instruct costs about 267x more per input token than DeepSeek Flash Latest.
How much does Qwen2.5 VL 72B Instruct cost per 1M tokens?
Qwen2.5 VL 72B Instruct costs $0.80 per 1M input tokens and $1 per 1M output tokens on the Qwen API, with a 128K-token context window.
How much does DeepSeek Flash Latest cost per 1M tokens?
DeepSeek Flash Latest costs $0.00 per 1M input tokens and $2.40 per 1M output tokens on the ~deepseek API, with a 1.0M-token context window.
What does a typical request cost on Qwen2.5 VL 72B Instruct vs DeepSeek Flash Latest?
For a typical request with 2,000 input and 500 output tokens: Qwen2.5 VL 72B Instruct costs $0.0021, DeepSeek Flash Latest costs $0.0012. At 1,000 requests/day for a month that is $63.00 for Qwen2.5 VL 72B Instruct vs $36.18 for DeepSeek Flash Latest.
Estimate your own workload with the per-model calculators (Qwen2.5 VL 72B Instruct, DeepSeek Flash Latest), browse all current prices on the live model pricing dashboard, or run these models head-to-head on a real prompt (free account required). Prices may differ from provider list prices for batch or tiered usage.
Pricing data via the OpenRouter models API.