Google model
Gemini 3.1 Flash Lite Preview API cost calculator
Gemini 3.1 Flash Lite Preview costs $0.25 per 1M input tokens and $1.50 per 1M output tokens on the Google API, with a 1.0M-token context window. A typical request (2K input, 500 output tokens) costs $0.0013 — about $37.50/month at 1,000 requests per day. Use the calculator below to model your exact workload.
Find me a cheaper modelWhat real workloads cost on Gemini 3.1 Flash Lite Preview
| Workload | Tokens (in / out) | Per request | Per 1K requests |
|---|---|---|---|
| Chatbot message | 500 / 300 | $0.0006 | $0.57 |
| RAG query with context | 4,000 / 500 | $0.0018 | $1.75 |
| Document summarization | 20,000 / 1,000 | $0.0065 | $6.50 |
| Agent coding session | 100,000 / 5,000 | $0.03 | $32.50 |
Gemini 3.1 Flash Lite Preview vs other Google models
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| Gemini 3.1 Flash Lite Preview | $0.25 | $1.50 | 1.0M |
| Gemini 3.8 Flash | $0.75 | $3.75 | 1.0M |
| Gemini 3.7 Flash | $0.75 | $3.75 | 1.0M |
| Gemini 3.6 Flash | $0.75 | $3.75 | 1.0M |
| Gemini 3.5 Flash Lite | $0.30 | $2.50 | 1.0M |
Frequently asked questions
How much does Gemini 3.1 Flash Lite Preview cost per 1M tokens?
According to current pricing data, Gemini 3.1 Flash Lite Preview costs $0.25 per 1M input tokens and $1.50 per 1M output tokens. Processing 1M tokens each way costs $1.75.
What does a typical API request to Gemini 3.1 Flash Lite Preview cost?
A typical request with 2,000 input tokens and 500 output tokens costs $0.0013. At 1,000 requests per day that is $1.25 daily, or about $37.50 per month.
What is Gemini 3.1 Flash Lite Preview's context window?
Gemini 3.1 Flash Lite Preview supports a 1.0M-token context window (1,048,576 tokens).
Does Gemini 3.1 Flash Lite Preview support prompt caching?
Yes. Cached input tokens cost $0.02 per 1M — a 90% discount versus the standard input rate, which matters for agents and chat apps that resend the same system prompt.
What is a cheaper alternative to Gemini 3.1 Flash Lite Preview?
DeepSeek Flash Latest from ~deepseek is currently the cheapest comparable option at $0.00 per 1M input tokens versus $0.25 for Gemini 3.1 Flash Lite Preview — roughly 83x cheaper on input.
Prices may differ from Google's official pricing page for batch, tiered, or fine-tuned usage. Compare all current prices on the live model pricing dashboard, see Gemini 3.1 Flash Lite Preview vs DeepSeek Flash Latest pricing, or run Gemini 3.1 Flash Lite Preview head-to-head against other models in the side-by-side comparison playground.
Pricing data via the OpenRouter models API.