Input
- Official Google
- $0.1
- Here (−50%)
- $0.05
Gemini 2.5 Flash-Lite is the cheapest Gemini model — built for massive-volume, latency-sensitive work like classification, extraction and routing.
| Rate | Official Google | Here (−50%) |
|---|---|---|
| Input | $0.1 | $0.05 |
| Cached input | $0.01 | $0.005 |
| Output | $0.4 | $0.2 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1M tokens. Max output: 64K tokens.
Create a free account, generate one key, and point any Gemini-compatible tool at https://gemini.api.apitoken.sale with model ID gemini-2.5-flash-lite — the native /v1beta/models/gemini-2.5-flash-lite:generateContent surface works, authenticated with x-goog-api-key. New accounts include $10 of API usage at official prices — enough to test the model before topping up.
Officially $0.10 per 1M input tokens and $0.40 per 1M output tokens, with cached input at $0.01. With the flat 50% apiToken.sale discount that is $0.05/$0.20 — the cheapest way to run Gemini.
High-volume, low-latency work: classification, extraction, summarization, routing and simple chat. For complex reasoning, step up to 2.5 Flash or the Gemini 3 line.
gemini-2.5-flash-lite. It works on the same apiToken.sale key and balance as every other supported Claude, GPT and Gemini model.
Run Gemini 2.5 Flash-Lite on the native Gemini API at a flat 50% off — instant key, prepaid balance, card or crypto.