Input
- Official OpenAI
- $0.2
- Here (−50%)
- $0.1
GPT-5.6 Luna is the fast, economical tier of the GPT-5.6 line — built for high-volume, latency-sensitive work at a fraction of the flagship price.
| Rate | Official OpenAI | Here (−50%) |
|---|---|---|
| Input | $0.2 | $0.1 |
| Cached input | $0.02 | $0.01 |
| Cache write | $0.25 | $0.125 |
| Output | $1.2 | $0.6 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 1.05M tokens. Max output: 128K tokens. Reasoning efforts: none, low, medium, high, xhigh, max.
Create a free account, generate one key, and point any OpenAI-compatible tool at https://router.apitoken.sale/v1 with model ID gpt-5.6-luna — Responses and Chat Completions both work, authenticated with Authorization: Bearer. Eligible new accounts include $5 of platform bonus credit — enough to test the model before topping up. Existing integrations on the legacy host https://openai.api.apitoken.sale/v1 keep working.
Officially $0.20 per 1M input tokens and $1.20 per 1M output tokens, with cached input at $0.02. With the flat 50% apiToken.sale discount that is $0.10/$0.60 — the cheapest way to run GPT-5.6.
High-volume, low-latency work: classification, extraction, summarization, routing and simple chat. For complex reasoning, step up to Terra or Sol.
gpt-5.6-luna. It works on the same apiToken.sale key, balance and OpenAI-compatible endpoint as every other GPT model.
Run GPT-5.6 Luna on the OpenAI-compatible API at a flat 50% off — instant key, prepaid balance, card or crypto.