Qwen3 30B A3B vs Granite 4.0 Micro: pricing comparison
Granite 4.0 Micro is the cheaper option at $0.02/$0.11 per 1M input/output tokens — Qwen3 30B A3B ($0.12/$0.50) costs about 7.1x more per input token. Full spec-by-spec breakdown below.
| Qwen3 30B A3B | Granite 4.0 Micro | |
|---|---|---|
| Input /1M tokens | $0.12 | $0.02 |
| Output /1M tokens | $0.50 | $0.11 |
| Cache read /1M | — | — |
| Context window | 131K | 131K |
| Provider | Qwen | Ibm-granite |
| Vision input | No | No |
| Released | Apr 2025 | Oct 2025 |
Real workload costs: Qwen3 30B A3B vs Granite 4.0 Micro
Cost per single request at common token profiles — the cheapest model for each workload is highlighted.
| Workload | Tokens (in / out) | Qwen3 30B A3B | Granite 4.0 Micro |
|---|---|---|---|
| Chatbot message | 500 / 300 | $0.0002 | <$0.0001 |
| RAG query with context | 4,000 / 500 | $0.0007 | $0.0001 |
| Document summarization | 20,000 / 1,000 | $0.0029 | $0.0005 |
| Agent coding session | 100,000 / 5,000 | $0.01 | $0.0023 |
Frequently asked questions
Which is cheaper: Qwen3 30B A3B vs Granite 4.0 Micro?
Granite 4.0 Micro is the cheapest on input tokens at $0.02 per 1M, and Granite 4.0 Micro is cheapest on output at $0.11 per 1M. Qwen3 30B A3B costs about 7.1x more per input token than Granite 4.0 Micro.
How much does Qwen3 30B A3B cost per 1M tokens?
Qwen3 30B A3B costs $0.12 per 1M input tokens and $0.50 per 1M output tokens on the Qwen API, with a 131K-token context window.
How much does Granite 4.0 Micro cost per 1M tokens?
Granite 4.0 Micro costs $0.02 per 1M input tokens and $0.11 per 1M output tokens on the Ibm-granite API, with a 131K-token context window.
What does a typical request cost on Qwen3 30B A3B vs Granite 4.0 Micro?
For a typical request with 2,000 input and 500 output tokens: Qwen3 30B A3B costs $0.0005, Granite 4.0 Micro costs <$0.0001. At 1,000 requests/day for a month that is $14.70 for Qwen3 30B A3B vs $2.70 for Granite 4.0 Micro.
Estimate your own workload with the per-model calculators (Qwen3 30B A3B, Granite 4.0 Micro), browse all current prices on the live model pricing dashboard, or run these models head-to-head on a real prompt (free account required). Prices may differ from provider list prices for batch or tiered usage.
Pricing data via the OpenRouter models API.