Models — Claude, GPT and Gemini
Current Claude, GPT and Gemini models with verified context windows and transparent reference pricing — one key, one balance.
Official rates behind every spend calculation
These official Anthropic, OpenAI and Google list rates calculate official API spend. B2C accounts pay 50% of that spend on every request; B2B rates are negotiated.
Claude · Anthropic Messages API
| Model | Context | Input / 1M | Output / 1M | Best for |
|---|---|---|---|---|
Claude Fable 5.1Latestclaude-fable-5-1 | 1M | $10 | $50 | Current Mythos GA — cheapest Fable cache reads |
Claude Opus 5claude-opus-5 | 1M | $5 | $25 | The new default for agentic coding |
Claude Fable 5claude-fable-5 | 1M | $10 | $50 | Mythos class — the longest-horizon runs |
Claude Opus 4.8claude-opus-4-8 | 1M | $5 | $25 | Deep agentic coding, hardest reasoning |
Claude Opus 4.7claude-opus-4-7 | 1M | $5 | $25 | Complex coding and long sessions |
Claude Sonnet 5claude-sonnet-5 | 1M | $2 | $10 | Balanced speed and quality, daily driver |
Claude Sonnet 4.6claude-sonnet-4-6 | 200K | $3 | $15 | Balanced speed and quality, daily driver |
Claude Haiku 4.5claude-haiku-4-5 | 200K | $1 | $5 | Fast, cheap, high-volume calls |
* Claude Fable 5.1 prompt-cache reads are $0.25 per 1M (0.025× input). Claude Fable 5 cache reads stay $1 per 1M.
GPT · OpenAI-compatible API
| Model | Context | Input / 1M | Output / 1M | Best for |
|---|---|---|---|---|
GPT-6 AstraLatestgpt-6-astra | 872K | $10 | $50 | Advanced reasoning, coding and image understanding |
GPT-5.6 Solgpt-5.6-sol | 1.05M | $4 | $20 | Flagship reasoning and agentic coding |
GPT-5.6 Terragpt-5.6-terra | 1.05M | $2 | $12 | Balanced daily driver at 40% of the flagship price |
GPT-5.6 Lunagpt-5.6-luna | 1.05M | $0.20 | $1.20 | Fast, cheap, high-volume calls |
GPT-5.5gpt-5.5 | 1.05M | $5 | $30 | Previous-generation flagship |
GPT-5.4gpt-5.4 | 1.05M | $2.50 | $15 | Proven balanced tier |
GPT rows are official OpenAI rates. GPT-6 Astra is the latest GPT model, with 872K maximum Codex context and 128K maximum output. GPT-5.6 Sol promotional pricing is $4 / $20 per 1M through 2026-11-21 and returns to $5 / $30 on 2026-11-22 UTC. gpt-5.6 is an alias of gpt-5.6-sol. Requests above 272K input tokens bill at 2× input and 1.5× output on the whole request.
Gemini · Google Gemini API
| Model | Context | Input / 1M | Output / 1M | Best for |
|---|---|---|---|---|
Gemini 3.8 FlashLatestgemini-3.8-flash | 1M | $0.75* | $3.75* | Current GA Flash — text, media, tools and SSE |
Gemini 3.7 Flashgemini-3.7-flash | 1M | $0.75* | $3.75* | Previous GA Flash — live-proven text and SSE |
Gemini 3.6 Flashgemini-3.6-flash | 1M | $0.75* | $3.75* | Previous Flash generation with broad controls |
Gemini 3.5 Flashgemini-3.5-flash | 1M | $1.50 | $9.00 | Proven high-throughput Flash |
Gemini 3 Flash Previewgemini-3-flash-preview | 1M | $0.50 | $3.00 | Cost-efficient Gemini 3 Flash preview |
Gemini 3.1 Pro Previewgemini-3.1-pro-preview | 1M | $2* | $12* | Pro-tier hardest reasoning, long context |
Gemini 3.1 Flash-Litegemini-3.1-flash-lite | 1M | $0.25 | $1.50 | Cheap, fast, high-volume calls |
Gemini 2.5 Flashgemini-2.5-flash | 1M | $0.30 | $2.50 | Proven previous-generation Flash |
Gemini 2.5 Flash-Litegemini-2.5-flash-lite | 1M | $0.10 | $0.40 | The cheapest Gemini model |
* Gemini 3.6 Flash, Gemini 3.7 Flash and Gemini 3.8 Flash promotional rates are $0.75 / $3.75 per 1M through 2026-12-31 and become $1.50 / $7.50 on 2027-01-01. The table resolves the effective rate at build time. Gemini 3.1 Pro Preview bills $4 / $18 per 1M above 200K input tokens. Gemini 3.1 Flash Image bills image output at $60 per 1M image-output tokens.