Qwen model
Qwen3 VL 30B A3B Thinking API cost calculator
Qwen3 VL 30B A3B Thinking costs $0.20 per 1M input tokens and $2.40 per 1M output tokens on the Qwen API, with a 262K-token context window. A typical request (2K input, 500 output tokens) costs $0.0016 — about $48.00/month at 1,000 requests per day. Use the calculator below to model your exact workload.
Find me a cheaper modelWhat real workloads cost on Qwen3 VL 30B A3B Thinking
| Workload | Tokens (in / out) | Per request | Per 1K requests |
|---|---|---|---|
| Chatbot message | 500 / 300 | $0.0008 | $0.82 |
| RAG query with context | 4,000 / 500 | $0.0020 | $2.00 |
| Document summarization | 20,000 / 1,000 | $0.0064 | $6.40 |
| Agent coding session | 100,000 / 5,000 | $0.03 | $32.00 |
Qwen3 VL 30B A3B Thinking vs other Qwen models
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| Qwen3 VL 30B A3B Thinking | $0.20 | $2.40 | 262K |
| Qwen3.8 Max (0902) | $2 | $6 | 1M |
| Qwen3.8 Flash | $0.15 | $0.47 | 1M |
| Qwen3.8 27B | $0.21 | $2.55 | 1M |
| Qwen3.8 2.4T A95B | $2 | $6 | 1.0M |
Frequently asked questions
How much does Qwen3 VL 30B A3B Thinking cost per 1M tokens?
According to current pricing data, Qwen3 VL 30B A3B Thinking costs $0.20 per 1M input tokens and $2.40 per 1M output tokens. Processing 1M tokens each way costs $2.60.
What does a typical API request to Qwen3 VL 30B A3B Thinking cost?
A typical request with 2,000 input tokens and 500 output tokens costs $0.0016. At 1,000 requests per day that is $1.60 daily, or about $48.00 per month.
What is Qwen3 VL 30B A3B Thinking's context window?
Qwen3 VL 30B A3B Thinking supports a 262K-token context window (262,144 tokens).
Does Qwen3 VL 30B A3B Thinking support prompt caching?
No cache-read pricing is published for Qwen3 VL 30B A3B Thinking, so every input token is billed at the full $0.20 per 1M rate.
What is a cheaper alternative to Qwen3 VL 30B A3B Thinking?
Granite 4.0 Micro from Ibm-granite is currently the cheapest comparable option at $0.02 per 1M input tokens versus $0.20 for Qwen3 VL 30B A3B Thinking — roughly 12x cheaper on input.
Compare all current prices on the live model pricing dashboard, see Qwen3 VL 30B A3B Thinking vs Granite 4.0 Micro pricing, or run Qwen3 VL 30B A3B Thinking head-to-head against other models in the side-by-side comparison playground.
Pricing data via the OpenRouter models API.