Input
- Official Google
- $1.5
- Here (−50%)
- $0.75
Gemini 3.6 Flash is the newest Gemini model — Google's frontier-class Flash for agentic coding, multimodal work and long-context tasks, at Flash-tier pricing.
| Rate | Official Google | Here (−50%) |
|---|---|---|
| Input | $1.5 | $0.75 |
| Cached input | $0.15 | $0.075 |
| Output | $7.5 | $3.75 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1M tokens. Max output: 64K tokens.
Create a free account, generate one key, and point any Gemini-compatible tool at https://gemini.api.apitoken.sale with model ID gemini-3.6-flash — the native /v1beta/models/gemini-3.6-flash:generateContent surface works, authenticated with x-goog-api-key. New accounts include $10 of API usage at official prices — enough to test the model before topping up.
Officially $1.50 per 1M input tokens and $7.50 per 1M output tokens, with cached input at $0.15. On apiToken.sale the same requests cost 50% less — $0.75/$3.75 at the flat discount applied to every call.
gemini-3.6-flash. Use it unchanged with the Google GenAI SDK or any Gemini-compatible tool pointed at https://gemini.api.apitoken.sale, with the key sent as x-goog-api-key.
3.6 Flash is newer and cheaper on output — $7.50 vs $9.00 per 1M at the same input price — so new projects should default to it. Keep 3.5 Flash only where prompts and evals are pinned to it.
Run Gemini 3.6 Flash on the native Gemini API at a flat 50% off — instant key, prepaid balance, card or crypto.