Input
- Official Google
- $0.3
- Here (−50%)
- $0.15
Gemini 2.5 Flash is the proven previous-generation Flash — a stable workhorse for production pipelines evaluated against the 2.5 line.
| Rate | Official Google | Here (−50%) |
|---|---|---|
| Input | $0.3 | $0.15 |
| Cached input | $0.03 | $0.015 |
| Output | $2.5 | $1.25 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1M tokens. Max output: 64K tokens.
Create a free account, generate one key, and point any Gemini-compatible tool at https://gemini.api.apitoken.sale with model ID gemini-2.5-flash — the native /v1beta/models/gemini-2.5-flash:generateContent surface works, authenticated with x-goog-api-key. New accounts include $10 of API usage at official prices — enough to test the model before topping up.
Officially $0.30 per 1M input tokens and $2.50 per 1M output tokens, with cached input at $0.03. With the flat 50% apiToken.sale discount that is $0.15/$1.25.
gemini-2.5-flash. Use it unchanged with the Google GenAI SDK or any Gemini-compatible tool pointed at https://gemini.api.apitoken.sale, with the key sent as x-goog-api-key.
2.5 Flash is far cheaper per token and proven in production; 3.5 Flash is the stronger current model. Keep 2.5 Flash where prompts and evals are pinned to it, default to the 3.x line for new work.
Run Gemini 2.5 Flash on the native Gemini API at a flat 50% off — instant key, prepaid balance, card or crypto.