Built for speed
OpenAI calls Flare its fastest model for high-quality, everyday images. It reports 50% lower latency than GPT Image 2 at comparable image quality.
GPT Image 2.5 Flare API: OpenAI's fast image model at $0.02 per image, any size.
gpt-image-2.5-flarefrom openai import OpenAIclient = OpenAI( base_url="https://api.zurelay.com/v1", api_key="YOUR_ZURELAY_KEY",)image = client.images.generate( model="gpt-image-2.5-flare", prompt="A red fox sitting in fresh snow at dawn, photo", size="1536x1024", quality="high",)print(image.data[0].url)Base URL https://api.zurelay.com/v1. Works with any OpenAI SDK.
Examples
Every image here is a single API call with the prompt underneath it, not retouched or upscaled. They're shown as compressed WebP; the API returns full-quality PNG.
Pricing
Pay as you go from prepaid credit, with no subscription and no minimum. Requests that fail are never billed.
| Rate | zurelay | OpenAI list | You save |
|---|---|---|---|
1K image 1024 × 1024, any quality | $0.02 | 62%off | |
2K image 2048 × 2048, any quality | $0.02 | ≈ | 91%off |
4K image 3840 × 2160, any quality | $0.02 | ≈ | 95%off |
One price at every size, always high quality. OpenAI bills images by output tokens, which grow with the image's area, so its 2K and 4K prices (≈) are worked out from its 1024 × 1024 price.
Move the sliders to your monthly usage.
$392.40 a year
Overview
GPT Image 2.5 Flare is OpenAI's fastest model for high-quality, everyday image generation, released on September 8, 2026. The GPT Image 2.5 Flare API on zurelay turns a text prompt into images through one OpenAI-compatible endpoint. Every image is made at high quality, at one flat price for any size.
GPT Image 2.5 Flare is one of two image models OpenAI released on September 8, 2026, next to GPT Image 2.5 Sunburst. OpenAI calls Flare its fastest model for high-quality, everyday image generation and positions it as the default for most applications. It says Flare cuts latency by 50% compared with GPT Image 2, with image quality comparable to that model. The GPT Image 2.5 Flare API on zurelay uses the model ID gpt-image-2.5-flare and also accepts the shorter name gpt-image-2.5.
Flare's recommended sizes are 1024x1024, 1536x1024 and 1024x1536, and it also takes custom WIDTHxHEIGHT sizes. Both edges must be multiples of 16, the aspect ratio must sit between 1:3 and 3:1, no edge may pass 3840 pixels, and the total must fall between 655,360 and 8,294,400 pixels. That reaches 4K frames like 3840x2160, but OpenAI treats anything above 2560x1440 as experimental. Quality can be low, medium, high, xhigh, max or auto, and auto is the default.
Flare can put text inside images. For exact wording, OpenAI advises quoting the text in your prompt, naming its position and typography, and checking spelling in the result, because precise text placement can still slip. For transparent assets, set background to transparent and output_format to png or webp. JPEG output is also available and returns faster than PNG.
OpenAI bills GPT Image 2.5 by tokens for text input, image input and image output. Higher quality and larger sizes both add output tokens, so OpenAI's cost per image climbs as you raise either one. zurelay makes every image at high quality and charges one flat price at every size, so a 4K image costs the same as a square. Send requests to POST https://api.zurelay.com/v1/images/generations in the OpenAI images format and get a URL back by default, or base64 with response_format set to b64_json.
Strengths
OpenAI calls Flare its fastest model for high-quality, everyday images. It reports 50% lower latency than GPT Image 2 at comparable image quality.
zurelay bills a flat price per image and makes every image at high quality. Larger sizes don't raise your cost.
Use any size with edges in multiples of 16 and a ratio from 1:3 to 3:1, up to 3840 pixels on the long edge.
Set background to transparent with PNG or WebP output to get cut-out images ready to drop into a layout.
Use cases
Post images, thumbnails and ad variations at the pace a content calendar needs.
Generate images inside your product, where users wait on the result and latency matters.
Draft many concepts, then rerun the best ones at a larger size. On zurelay every run costs the same.
Catalog images, placeholders and variation sets where a flat price per image keeps the bill predictable.
Get started
No waitlist and no new SDK. If your code already talks to OpenAI, it already talks to zurelay.
Sign up, add credit and create an API key. Set a monthly budget or a rate limit per key if you like.
Change the base URL. Everything else in your code stays the same.
https://api.zurelay.com/v1Send a prompt to /v1/images/generations with a size and aspect ratio, and get back a link to your image.
gpt-image-2.5-flareWorks with the tools you already use
Compare
$0.02 per image
Pick Sunburst when fine detail and polish matter more than speed, since OpenAI tunes it for image quality.
GPT Image 2.5 Sunburst APIfrom $0.035 per image
Pick Nano Banana Pro to compare Google's image model, with fixed 1K, 2K and 4K outputs and a focus on legible text.
Nano Banana Pro API$2.50 / $12.50 per 1M tokens
Use GPT-6 Astra to write detailed image prompts or review finished images; it reads images but does not create them.
GPT-6 Astra APIFAQ
Anything else, write to sales@zurelay.com and an engineer will answer.
zurelay charges $0.02 per image for gpt-image-2.5-flare at every size, and every image is made at high quality. OpenAI bills by tokens, and its cost for a 1024x1024 image at high quality is about $0.0527, so you save 62%. OpenAI's cost rises with larger sizes, while zurelay's stays flat.
Yes. zurelay serves GPT Image 2.5 Flare at $0.02 per image, 62% below OpenAI's approximate cost for a 1024x1024 high-quality image. The gap grows at larger sizes, because zurelay's price does not change with size.
Create an OpenAI client with the base URL set to https://api.zurelay.com/v1 and your zurelay API key. Call images.generate with model set to gpt-image-2.5-flare, a prompt, a size such as 1536x1024. You can also pass background, output_format and n (1 to 4). The response holds an image URL by default; set response_format to b64_json to get base64 instead.
The recommended sizes are 1024x1024, 1536x1024 and 1024x1536. Custom sizes work when both edges are multiples of 16, the ratio is between 1:3 and 3:1, no edge passes 3840 pixels and the total is between 655,360 and 8,294,400 pixels. OpenAI treats sizes above 2560x1440 as experimental. On zurelay every image is made at high quality, whatever quality you pass, and a 1024x1024 square comes back at 1254x1254.
Start with Flare. OpenAI calls it its fastest model for everyday image generation and the default for most applications. Switch to Sunburst when you need the most detail and precision and can wait longer per image. Both cost the same flat $0.02 per image on zurelay, so test your prompts on each.
Yes to both. For exact text, put the words in quotes, say where they go and what typography to use, then check spelling, since OpenAI notes precise text placement can still slip. For a transparent background, set background to transparent and output_format to png or webp.
Images are stored for 30 days behind the URL in the response. Download any image you want to keep before then. If you request b64_json, the image data comes back in the response itself.
Yes. zurelay runs the same gpt-image-2.5-flare model and accepts requests in the OpenAI images format. Image generation varies from run to run, so the same prompt will not return byte-identical images. zurelay is an independent service and is not affiliated with OpenAI.
No. zurelay offers text-to-image generation only, through /v1/images/generations. Image edits, inpainting and reference-image inputs are not available.
Create a key in seconds, point the SDK you already use at zurelay, and every request costs up to 85% less from the first token.