Inclusionai model
Ling 3.1 Flash API cost calculator
Ling 3.1 Flash costs free per 1M input tokens and free per 1M output tokens on the Inclusionai API, with a 262K-token context window. A typical request (2K input, 500 output tokens) costs <$0.0001 — about <$0.0001/month at 1,000 requests per day. Use the calculator below to model your exact workload.
Find me a cheaper modelWhat real workloads cost on Ling 3.1 Flash
| Workload | Tokens (in / out) | Per request | Per 1K requests |
|---|---|---|---|
| Chatbot message | 500 / 300 | <$0.0001 | <$0.0001 |
| RAG query with context | 4,000 / 500 | <$0.0001 | <$0.0001 |
| Document summarization | 20,000 / 1,000 | <$0.0001 | <$0.0001 |
| Agent coding session | 100,000 / 5,000 | <$0.0001 | <$0.0001 |
Ling 3.1 Flash vs other Inclusionai models
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| Ling 3.1 Flash | free | free | 262K |
| Ling 3.0 Flash VL | $0.02 | $0.06 | 262K |
| Ling 3.0 Flash Fin | $0.04 | $0.12 | 262K |
| Ling 3.0 Flash | $0.02 | $0.06 | 262K |
Frequently asked questions
How much does Ling 3.1 Flash cost per 1M tokens?
According to current pricing data, Ling 3.1 Flash costs free per 1M input tokens and free per 1M output tokens. Processing 1M tokens each way costs <$0.0001.
What does a typical API request to Ling 3.1 Flash cost?
A typical request with 2,000 input tokens and 500 output tokens costs <$0.0001. At 1,000 requests per day that is <$0.0001 daily, or about <$0.0001 per month.
What is Ling 3.1 Flash's context window?
Ling 3.1 Flash supports a 262K-token context window (262,144 tokens).
Does Ling 3.1 Flash support prompt caching?
No cache-read pricing is published for Ling 3.1 Flash, so every input token is billed at the full free per 1M rate.
Compare all current prices on the live model pricing dashboard, or run Ling 3.1 Flash head-to-head against other models in the side-by-side comparison playground.
Pricing data via the OpenRouter models API.