OpenAI: GPT Image 2.5 Flare

openai/gpt-image-2.5-flare

GPT Image 2.5 Flare is an image generation and editing model from OpenAI, positioned as the speed-oriented tier of the GPT Image 2.5 series. It is suited to high-volume everyday generation, creator co...

vision
MODALITIES
INPUT PRICE
OUTPUT PRICE
CONTEXT
400K
RELEASED
Sep 9, 2026
ProviderCacheUptimeChat
5.00(1)$0.01/call
Hit rate0.0%
Cache read$0.5

Capabilities

Input modalities
imagetext
Output modalities
image
Features
frequency_penaltylogit_biaslogprobsmax_tokenspresence_penaltyresponse_formatseedstopstructured_outputstemperaturetop_logprobstop_p
1

Get API Key

Create an API key from the Keys page, then set it as an environment variable:

export ONLIST_API_KEY=sk-...
2

Make your first request

Endpoints

POSThttps://onlist.io/v1/images/generations
Request Headers
Authorization:Bearer $ONLIST_API_KEY
Content-Type:application/json
Model:openai/gpt-image-2.5-flare

Code samples

curl https://onlist.io/v1/images/generations \
  -H "Authorization: Bearer $ONLIST_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
       "model": "openai/gpt-image-2.5-flare",
       "prompt": "A white siamese cat",
       "n": 1,
       "size": "1024x1024"
     }'

Replace $ONLIST_API_KEY with the API key from your Keys page.

Authentication

All requests must include an Authorization: Bearer <TOKEN> header. Generate keys from the Keys page; keys can be scoped to specific models, groups, IP ranges, and rate limits.

Supported parameters

NameTypeDescription
frequency_penaltynumberPenalizes new tokens based on their existing frequency in the text (-2.0 to 2.0).
logit_bias
logprobsbooleanWhether to return log probabilities of the output tokens.
max_tokensintegerMaximum number of tokens to generate in the completion.
presence_penaltynumberPenalizes new tokens based on whether they appear in the text so far (-2.0 to 2.0).
response_formatobjectSpecifies the output format. Use {"type": "json_object"} for JSON mode.
seedintegerIf specified, the system will attempt deterministic sampling for reproducible results.
stopstring | arrayUp to 4 sequences where the API will stop generating tokens.
structured_outputs
temperaturenumberSampling temperature between 0 and 2. Higher values make output more random.
top_logprobsintegerNumber of most likely tokens to return at each position (0-20). Requires logprobs: true.
top_pnumberNucleus sampling. The model considers tokens with top_p probability mass.

These are the request parameters this model accepts. Parameter semantics follow the OpenAI Chat Completions specification.