Input
- Official Google
- $0.75
- Here (−50%)
- $0.375
Gemini 3.8 Flash is Google's current GA Flash model, available through the native Gemini API with streaming and authoritative token usage.
| Rate | Official Google | Here (−50%) |
|---|---|---|
| Input | $0.75 | $0.375 |
| Cached input | $0.075 | $0.0375 |
| Output | $3.75 | $1.875 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1M tokens. Max output: 64K tokens.
Create a free account, generate one key, and point any Gemini-compatible tool at https://router.apitoken.sale with model ID gemini-3.8-flash — the native /v1beta/models/gemini-3.8-flash:generateContent surface works, authenticated with x-goog-api-key. Eligible new accounts include $5 of platform bonus credit — enough to test the model before topping up. Existing integrations on the legacy host https://gemini.api.apitoken.sale keep working.
Always use gemini-3.8-flash. apiToken.sale keeps any private upstream routing name internal and returns only this public ID.
Through 2026 the official promotional rate is $0.75 per 1M input tokens, $0.075 cached input and $3.75 output. The flat 50% apiToken.sale B2C discount makes that $0.375, $0.0375 and $1.875.
Text, image, audio (inline WAV), video (inline MP4) and PDF input, function calling, JSON structured output, implicit prompt caching, Google Search grounding, countTokens and incremental SSE streaming, plus explicit thinking levels low, medium and high through reasoning_effort.
Run Gemini 3.8 Flash on the native Gemini API at a flat 50% off — instant key, prepaid balance, card or crypto.