Models

Models — Claude, GPT and Gemini

Current Claude, GPT and Gemini models with verified context windows and transparent reference pricing — one key, one balance.

Official list rates

Official rates behind every spend calculation

These official Anthropic, OpenAI and Google list rates calculate official API spend. B2C accounts pay 50% of that spend on every request; B2B rates are negotiated.

Claude · Anthropic Messages API

ModelContextInput / 1MOutput / 1MBest for
Claude Fable 5.1Latest
claude-fable-5-1
1M$10$50Current Mythos GA — cheapest Fable cache reads
Claude Opus 5
claude-opus-5
1M$5$25The new default for agentic coding
Claude Fable 5
claude-fable-5
1M$10$50Mythos class — the longest-horizon runs
Claude Opus 4.8
claude-opus-4-8
1M$5$25Deep agentic coding, hardest reasoning
Claude Opus 4.7
claude-opus-4-7
1M$5$25Complex coding and long sessions
Claude Sonnet 5
claude-sonnet-5
1M$2$10Balanced speed and quality, daily driver
Claude Sonnet 4.6
claude-sonnet-4-6
200K$3$15Balanced speed and quality, daily driver
Claude Haiku 4.5
claude-haiku-4-5
200K$1$5Fast, cheap, high-volume calls

* Claude Fable 5.1 prompt-cache reads are $0.25 per 1M (0.025× input). Claude Fable 5 cache reads stay $1 per 1M.

GPT · OpenAI-compatible API

ModelContextInput / 1MOutput / 1MBest for
GPT-6 AstraLatest
gpt-6-astra
872K$10$50Advanced reasoning, coding and image understanding
GPT-5.6 Sol
gpt-5.6-sol
1.05M$4$20Flagship reasoning and agentic coding
GPT-5.6 Terra
gpt-5.6-terra
1.05M$2$12Balanced daily driver at 40% of the flagship price
GPT-5.6 Luna
gpt-5.6-luna
1.05M$0.20$1.20Fast, cheap, high-volume calls
GPT-5.5
gpt-5.5
1.05M$5$30Previous-generation flagship
GPT-5.4
gpt-5.4
1.05M$2.50$15Proven balanced tier

GPT rows are official OpenAI rates. GPT-6 Astra is the latest GPT model, with 872K maximum Codex context and 128K maximum output. GPT-5.6 Sol promotional pricing is $4 / $20 per 1M through 2026-11-21 and returns to $5 / $30 on 2026-11-22 UTC. gpt-5.6 is an alias of gpt-5.6-sol. Requests above 272K input tokens bill at 2× input and 1.5× output on the whole request.

Gemini · Google Gemini API

ModelContextInput / 1MOutput / 1MBest for
Gemini 3.8 FlashLatest
gemini-3.8-flash
1M$0.75*$3.75*Current GA Flash — text, media, tools and SSE
Gemini 3.7 Flash
gemini-3.7-flash
1M$0.75*$3.75*Previous GA Flash — live-proven text and SSE
Gemini 3.6 Flash
gemini-3.6-flash
1M$0.75*$3.75*Previous Flash generation with broad controls
Gemini 3.5 Flash
gemini-3.5-flash
1M$1.50$9.00Proven high-throughput Flash
Gemini 3 Flash Preview
gemini-3-flash-preview
1M$0.50$3.00Cost-efficient Gemini 3 Flash preview
Gemini 3.1 Pro Preview
gemini-3.1-pro-preview
1M$2*$12*Pro-tier hardest reasoning, long context
Gemini 3.1 Flash-Lite
gemini-3.1-flash-lite
1M$0.25$1.50Cheap, fast, high-volume calls
Gemini 2.5 Flash
gemini-2.5-flash
1M$0.30$2.50Proven previous-generation Flash
Gemini 2.5 Flash-Lite
gemini-2.5-flash-lite
1M$0.10$0.40The cheapest Gemini model

* Gemini 3.6 Flash, Gemini 3.7 Flash and Gemini 3.8 Flash promotional rates are $0.75 / $3.75 per 1M through 2026-12-31 and become $1.50 / $7.50 on 2027-01-01. The table resolves the effective rate at build time. Gemini 3.1 Pro Preview bills $4 / $18 per 1M above 200K input tokens. Gemini 3.1 Flash Image bills image output at $60 per 1M image-output tokens.