v1

latestOpenAPI 3.1.02026-07-2611077199.5 KB

post/v1/images/generations

Request body

promptstring required

Text description of the image. gpt-image-1.5, gpt-image-1, and gpt-image-2 support up to 32,000 characters; dall-e-2 supports up to 1,000 characters; dall-e-3 supports up to 4,000 characters.

background'transparent' | 'opaque' | 'auto'

Sets the transparency of generated image backgrounds. Applies only to GPT Image models. Choose transparent, opaque, or auto (default). When set to transparent, output_format must support transparency; use png or webp.

model'gpt-image-2' | 'gpt-image-1.5' | 'gpt-image-1' | 'dall-e-2' | 'dall-e-3' required

The model to use for image generation.

moderation'low' | 'auto'

Controls the content moderation level for images generated by GPT Image models. Choose low for less restrictive filtering or auto (default).

ninteger

Number of images to generate, default 1. Available range 1-10; dall-e-3 is fixed at 1.

output_compressioninteger

Compression level for generated images (0-100). Applies only to GPT Image models with output_format set to webp or jpeg; defaults to 100.

output_format'png' | 'jpeg' | 'webp'

Return format for generated images. Applies only to GPT Image models. Choose png (default), jpeg, or webp.

partial_imagesinteger

Number of partial images to return in the streaming response, from 0 to 3; defaults to 0. Available only when stream is true.

qualitystring

Image quality. auto (default) automatically selects the best quality for the model.

GPT Image models support high, medium, and low;

dall-e-3 supports hd and standard;

dall-e-2 only supports standard.

response_format'url' | 'b64_json'

Return data format. Only dall-e-2 and dall-e-3 support url or b64_json; URLs are valid for 60 minutes. GPT Image models always return b64_json.

sizestring

Image size.

GPT Image models support 1024x1024, 1536x1024, 1024x1536, and auto (default).

gpt-image-2 support arbitrary WIDTHxHEIGHT resolution strings where both dimensions are divisible by 16 and the aspect ratio is between 1:3 and 3:1.

dall-e-2 supports 256x256, 512x512, 1024x1024;

dall-e-3 supports 1024x1024, 1792x1024, 1024x1792.

streamboolean

Whether to stream generation progress. When enabled, use partial_images to receive partial image events.

style'vivid' | 'natural'

Style of generated image. Applies only to dall-e-3.

userstring

A unique identifier representing your end-user, which can help OpenAI monitor and detect abuse.

Response

Request successful, returns image generation results (supports url/base64 formats)

createdinteger required

Generation timestamp (seconds)

Example response

{
  "created": 1745818538,
  "data": [
    {
      "url": "https://filesystem.site/cdn/20250428/5LsFMT0Def02RVVUkwJ4v7IslleMzo.webp",
      "revised_prompt": "A cute fluffy cat with big, bright eyes, soft fur, and a playful expression. It has a light gray coat with hints of white on its paws and tail. The cat is sitting in a cozy setting, surrounded by soft cushions, with its tail wrapped around its paws. Its ears are perked up, and it's looking directly at the viewer, exuding a sense of curiosity and warmth."
    }
  ],
  "usage": {
    "input_tokens": 9,
    "input_tokens_details": {
      "text_tokens": 9
    },
    "output_tokens": 4160,
    "total_tokens": 4169
  }
}