Qwen model
Qwen2.5 VL 72B Instruct API cost calculator
Qwen2.5 VL 72B Instruct costs $0.80 per 1M input tokens and $1 per 1M output tokens on the Qwen API, with a 128K-token context window. A typical request (2K input, 500 output tokens) costs $0.0021 — about $63.00/month at 1,000 requests per day. Use the calculator below to model your exact workload.
Find me a cheaper modelWhat real workloads cost on Qwen2.5 VL 72B Instruct
| Workload | Tokens (in / out) | Per request | Per 1K requests |
|---|---|---|---|
| Chatbot message | 500 / 300 | $0.0007 | $0.70 |
| RAG query with context | 4,000 / 500 | $0.0037 | $3.70 |
| Document summarization | 20,000 / 1,000 | $0.02 | $17.00 |
| Agent coding session | 100,000 / 5,000 | $0.09 | $85.00 |
Qwen2.5 VL 72B Instruct vs other Qwen models
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| Qwen2.5 VL 72B Instruct | $0.80 | $1 | 128K |
| Qwen3.8 Max (0902) | $2 | $6 | 1M |
| Qwen3.8 Flash | $0.15 | $0.47 | 1M |
| Qwen3.8 27B | $0.21 | $2.55 | 1M |
| Qwen3.8 2.4T A95B | $2 | $6 | 1.0M |
Frequently asked questions
How much does Qwen2.5 VL 72B Instruct cost per 1M tokens?
According to current pricing data, Qwen2.5 VL 72B Instruct costs $0.80 per 1M input tokens and $1 per 1M output tokens. Processing 1M tokens each way costs $1.80.
What does a typical API request to Qwen2.5 VL 72B Instruct cost?
A typical request with 2,000 input tokens and 500 output tokens costs $0.0021. At 1,000 requests per day that is $2.10 daily, or about $63.00 per month.
What is Qwen2.5 VL 72B Instruct's context window?
Qwen2.5 VL 72B Instruct supports a 128K-token context window (128,000 tokens).
Does Qwen2.5 VL 72B Instruct support prompt caching?
Yes. Cached input tokens cost $0.40 per 1M — a 50% discount versus the standard input rate, which matters for agents and chat apps that resend the same system prompt.
What is a cheaper alternative to Qwen2.5 VL 72B Instruct?
Granite 4.0 Micro from Ibm-granite is currently the cheapest comparable option at $0.02 per 1M input tokens versus $0.80 for Qwen2.5 VL 72B Instruct — roughly 47x cheaper on input.
Compare all current prices on the live model pricing dashboard, see Qwen2.5 VL 72B Instruct vs Granite 4.0 Micro pricing, or run Qwen2.5 VL 72B Instruct head-to-head against other models in the side-by-side comparison playground.
Pricing data via the OpenRouter models API.