Input
- Official Google
- $2
- Here (−50%)
- $1
Gemini 3.1 Pro Preview is Google's Pro-tier reasoning model — the strongest Gemini for hard reasoning and long-horizon agentic work, with long-context rates above 200K input tokens.
| Rate | Official Google | Here (−50%) |
|---|---|---|
| Input | $2 | $1 |
| Cached input | $0.2 | $0.1 |
| Output | $12 | $6 |
| Long-context input (>200K) | $4 | $2 |
| Long-context output (>200K) | $18 | $9 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1M tokens. Max output: 64K tokens.
Create a free account, generate one key, and point any Gemini-compatible tool at https://gemini.api.apitoken.sale with model ID gemini-3.1-pro-preview — the native /v1beta/models/gemini-3.1-pro-preview:generateContent surface works, authenticated with x-goog-api-key. New accounts include $10 of API usage at official prices — enough to test the model before topping up.
Officially $2 per 1M input tokens and $12 per 1M output tokens, with cached input at $0.20; above 200K input tokens the whole request bills at $4/$18. On apiToken.sale the flat 50% discount applies to every call — $1/$6, or $2/$9 at long-context rates.
gemini-3.1-pro-preview. Use it unchanged with the Google GenAI SDK or any Gemini-compatible tool pointed at https://gemini.api.apitoken.sale, with the key sent as x-goog-api-key.
3.6 Flash covers most workloads at a lower token price; route the hardest reasoning and longest-horizon runs to 3.1 Pro Preview. Both run on the same key, balance and endpoint.
Run Gemini 3.1 Pro Preview on the native Gemini API at a flat 50% off — instant key, prepaid balance, card or crypto.