GPT-Image-2.5 Flare
Default
Fast, high-quality everyday image generation
Fast, high-quality everyday image generation
Performance
Higher
Speed
Very fast
Price
$5•$30
Input•Output
Input
Text, image
Output
Image
GPT Image 2.5 Flare is our fastest model for high-quality, everyday image generation. It accepts text and image inputs and produces image outputs. It supports low, medium, high, xhigh, max, and auto quality settings. Select it directly in the Image API or as the model of the Responses API image generation tool. Learn more in the image generation guide.
Pricing
Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the pricing page.
Text tokens
Per 1M tokens
Input
$5.00
Cached input
$1.25
Image tokens
Per 1M tokens
Input
$8.00
Cached input
$2.00
Output
$30.00
Image output costs $30 per million tokens. Text output is not billed because this model outputs images, not text.
Token rates match GPT Image 2. The GPT Image 2 calculator does not estimate GPT Image 2.5 token consumption.
Modalities
Text
Input only
Image
Input and output
Audio
Not supported
Video
Not supported
Endpoints
Chat Completions
v1/chat/completions
Responses
v1/responses
Realtime
v1/realtime
Realtime translation
v1/realtime/translations
Realtime transcription
v1/realtime/transcription_sessions
Assistants
v1/assistants
Batch
v1/batch
Fine-tuning
v1/fine-tuning
Embeddings
v1/embeddings
Image generation
v1/images/generations
Videos
v1/videos
Image edit
v1/images/edits
Speech generation
v1/audio/speech
Transcription
v1/audio/transcriptions
Translation
v1/audio/translations
Moderation
v1/moderations
Completions (legacy)
v1/completions
Features
Streaming
Not supported
Function calling
Not supported
Structured outputs
Not supported
Fine-tuning
Not supported
Predicted outputs
Not supported
Model IDs
Use the undated model ID or pin the dated snapshot in your API requests.
gpt-image-2.5-flare
gpt-image-2.5-flare-2026-09-08
gpt-image-2.5-flare
gpt-image-2.5-flare-2026-09-08
Rate limits
Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.
See your organization's limits for the rate limits that apply to your account.