Input
- Official OpenAI
- $0.1
- Here (−50%)
- $0.05
GPT-6 Luna is the lightweight, low-cost GPT-6 model for high-volume and latency-sensitive work on the OpenAI-compatible endpoint. Choose reasoning effort from none to max, send images, and use structured outputs or function tools.
| Rate | Official OpenAI | Here (−50%) |
|---|---|---|
| Input | $0.1 | $0.05 |
| Cached input | $0.01 | $0.005 |
| Cache write | $0.125 | $0.0625 |
| Output | $0.5 | $0.25 |
Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: 872K tokens. Max output: 128K tokens. Reasoning efforts: none, low, medium, high, xhigh, max.
Create a free account, generate one key, and point any OpenAI-compatible tool at https://router.apitoken.sale/v1 with model ID gpt-6-luna — Responses and Chat Completions both work, authenticated with Authorization: Bearer. Eligible new accounts include $5 of platform bonus credit — enough to test the model before topping up. Existing integrations on the legacy host https://openai.api.apitoken.sale/v1 keep working.
Official rates per 1M tokens are $0.10 input, $0.01 cached input, $0.125 cache write and $0.50 output. With the normal 50% B2C discount, these are $0.05, $0.005, $0.0625 and $0.25. OpenKeys uses the official rates.
GPT-6 Luna is a separate, newer model with its own ID and lower official rates ($0.10/$0.50 against $0.20/$1.20). Both stay available; choose by model ID.
Set model to gpt-6-luna on https://router.apitoken.sale/v1. Supported reasoning levels are none, low, medium, high, xhigh and max. Use service_tier priority for Fast.
Run GPT-6 Luna on the OpenAI-compatible API at a flat 50% off — instant key, prepaid balance, card or crypto.