Input
- Official Google
- $0.25
- Here (−50%)
- $0.125
Gemini 3.1 Flash-Lite is the economical tier of the Gemini 3 line — built for high-volume, latency-sensitive work at a fraction of Flash pricing.
| Rate | Official Google | Here (−50%) |
|---|---|---|
| Input | $0.25 | $0.125 |
| Cached input | $0.025 | $0.013 |
| Output | $1.5 | $0.75 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1M tokens. Max output: 64K tokens.
Create a free account, generate one key, and point any Gemini-compatible tool at https://gemini.api.apitoken.sale with model ID gemini-3.1-flash-lite — the native /v1beta/models/gemini-3.1-flash-lite:generateContent surface works, authenticated with x-goog-api-key. New accounts include $10 of API usage at official prices — enough to test the model before topping up.
Officially $0.25 per 1M input tokens and $1.50 per 1M output tokens, with cached input at $0.025. With the flat 50% apiToken.sale discount that is $0.125/$0.75.
gemini-3.1-flash-lite. Point any Gemini-compatible client at https://gemini.api.apitoken.sale and send it as the model, with the key in x-goog-api-key.
Flash-Lite handles bulk, latency-sensitive work at a fraction of the price; step up to gemini-3.6-flash for agentic coding and harder reasoning. Many teams route by task on the same key.
Run Gemini 3.1 Flash-Lite on the native Gemini API at a flat 50% off — instant key, prepaid balance, card or crypto.