gpt-image-2.5-flare

GPT Image 2.5 Flare API — price per token

GPT Image 2.5 Flare is the faster everyday ChatGPT Images 2.5 model — text and reference images in, rendered images out, metered per token at the same rates as GPT Image 2 on the same prepaid balance.

Pricing per 1M tokens

RateOfficial OpenAIHere (−50%)
Input$5$2.5
Cached input$1.25$0.625
Cache write$0$0
Image output$30$15
Image input$8$4

Input

Official OpenAI
$5
Here (−50%)
$2.5

Cached input

Official OpenAI
$1.25
Here (−50%)
$0.625

Cache write

Official OpenAI
$0
Here (−50%)
$0

Image output

Official OpenAI
$30
Here (−50%)
$15

Image input

Official OpenAI
$8
Here (−50%)
$4

Every request is metered at the official rate first, then the flat 50% B2C discount is subtracted before it touches your prepaid balance. Context window: per request. Max output: 1 image.

Best for

  • Everyday product image generation from a prompt on a bounded low-cost profile.
  • Faster edits against one to five PNG, JPEG or WebP references.
  • Pipelines that already run GPT Image 2 here — Flare is a drop-in Images HTTP id.

Good to know

  • gpt-image-2.5-flare is the alias of the immutable snapshot gpt-image-2.5-flare-2026-09-08 — the same model at the same price. Official token rates match GPT Image 2.
  • Generation is POST /v1/images/generations, editing is POST /v1/images/edits with up to five PNG, JPEG or WebP references (each ≤50 MB) — on the unified router (router.apitoken.sale/v1) or the legacy OpenAI host. Official GPT Image fields are accepted and converted. Multipart mask is ignored. Region inpaint (and hosted image_generation from a GPT text model) is POST /v1/responses with input_image plus tools image_generation; Chat Completions maps the same tool. Responses forwards output_format jpeg/webp and partial_images 1..=3, rewrites background=transparent to opaque, and drops input_fidelity. Mask image_url is PNG data URL only; no file_id.
  • Images HTTP converts official client fields onto the Codex envelope (native size=auto, quality=low, background=opaque). Extra keys such as user, stream and moderation are ignored. quality medium/high/xhigh/max/auto still generate at low on a clean prompt. jpeg/webp transcode the native PNG locally. n=1..10 runs sequential native turns. size including 2048x2048 and 3840x2160 maps to 1:1 or 16:9; the PNG stays ~1.57 megapixels and the response size is the IHDR. background=transparent prepends a cutout sentence — real alpha is prompt-steered, not a JSON lock. Edits accept image or image[]; mask is ignored.
  • Image output bills per image-output token; cached text/image input bills at 25% of the fresh rate. Production low canary (2026-09-09): HTTP 200, RGB PNG 1254×1254, 515 image-output tokens.

How to use GPT Image 2.5 Flare

Create a free account, generate one key, and point any OpenAI-compatible tool at https://router.apitoken.sale/v1 with model ID gpt-image-2.5-flare — Responses and Chat Completions both work, authenticated with Authorization: Bearer. Eligible new accounts include $5 of platform bonus credit — enough to test the model before topping up. Existing integrations on the legacy host https://openai.api.apitoken.sale/v1 keep working.

Frequently asked questions

How much does the GPT Image 2.5 Flare API cost?

Officially $5 per 1M text input tokens, $8 per 1M image input tokens and $30 per 1M image output tokens (cached input at 25%). On apiToken.sale the same calls cost 50% less — $2.50/$4/$15 — at the flat discount applied to every call. Rates match GPT Image 2.

What is the model ID for GPT Image 2.5 Flare?

gpt-image-2.5-flare, an alias of the immutable snapshot gpt-image-2.5-flare-2026-09-08. Send it as the model field of POST /v1/images/generations or /v1/images/edits on https://router.apitoken.sale/v1 (the legacy https://openai.api.apitoken.sale/v1 serves the same routes) with your Bearer key.

Does quality xhigh or max produce a higher-tier image?

No. On this ChatGPT pool, quality medium/high/xhigh/max/auto still generate at low on a clean prompt. A cutout/logo prompt can upgrade the native echo to medium and real RGBA even when JSON said opaque. Inspect the PNG; a mismatch is still HTTP 200.

Can I send an OpenAI Images mask file?

POST /v1/images/edits ignores the mask field so the SDK still gets a full-image edit. To inpaint a region, call POST /v1/responses with a GPT text model, the source PNG as input_image, and image_generation.input_image_mask.image_url as a PNG data URL. You can set output_format jpeg|webp and partial_images 1..=3. file_id is not supported. Chat Completions maps the same tool onto Responses.

Does the same balance really cover image generation?

Yes. GPT Image 2.5 Flare debits the same prepaid balance as every Claude, GPT and Gemini model on the account — no separate image plan or key.

Run GPT Image 2.5 Flare on the OpenAI-compatible API at a flat 50% off — instant key, prepaid balance, card or crypto.