GPT Image

GPT-Image-1

by OpenAI · Deprecated; currently accessible and scheduled to shut down on 2026-10-23

GPT-Image-1 is OpenAI’s API model for generating and editing images from text and image inputs. It supports reference images, masks, transparent backgrounds, multiple quality levels, several output sizes, and PNG, JPEG, or WebP output. The deprecated model is scheduled to shut down on October 23, 2026, making lifecycle planning essential for new integrations.

Image generation
GPT-Image-1 is OpenAI’s dedicated API model for creating and editing images rather than generating text. Released on April 23, 2025, it brought the image-generation capabilities associated with ChatGPT’s native image experience to developer applications. The model can create images from prompts, modify supplied images, use masks for targeted edits, and produce several output sizes and formats. It is currently deprecated, so new projects should consider migration planning before adopting it.
Outputs

What GPT-Image-1 can produce

Image generation
Inputs

What it can understand

Text Images Multimodal input
Capabilities

Supported features

Prompt caching Multimodal output
Model profile

Performance characteristics

6/10 Speed
5/10 Cost efficiency
Specifications

Technical details

Model family GPT Image
Model type Image Generation
Release date 2025-04-23
Status Deprecated; currently accessible and scheduled to shut down on 2026-10-23
Deprecation date 2026-06-02
Shutdown date 2026-10-23
Knowledge cutoff notes

OpenAI does not publish a direct knowledge-cutoff date for GPT-Image-1 in the cited model documentation. The model's image-generation behavior and image-editing capabilities should not be treated as evidence of a separately documented knowledge cutoff.

Model notes

Canonical model ID is gpt-image-1. The model accepts text and image inputs and produces image outputs. It supports image generation and editing, reference images, masks, transparent backgrounds, low/medium/high/auto quality, 1024x1024, 1024x1536, 1536x1024, and auto sizing, plus PNG, JPEG, and WebP output. OpenAI documents no streaming, function calling, structured outputs, or fine-tuning for this exact model. GPT-Image-1 is deprecated but remains accessible as of September 23, 2026; OpenAI's model-specific documentation states that it is scheduled to shut down on October 23, 2026. Per-image pricing does not include text and image token charges.

Cost

Model pricing

Input Text input: $5.00 per 1M tokens; cached text input: $1.25 per 1M tokens; image input: $10.00 per 1M image tokens; cached image input: $2.50 per 1M image tokens
Output Image generation per image: low $0.011 at 1024x1024 or $0.016 at 1024x1536 and 1536x1024; medium $0.042 or $0.063; high $0.167 or $0.25. Image output tokens: $40.00 per 1M tokens.
Model guide

GPT-Image-1: OpenAI’s API Model for Image Generation and Editing

GPT-Image-1 is OpenAI’s natively multimodal image-generation and editing model for API applications. It accepts text and image inputs, supports reference images and mask-based edits, and returns images in PNG, JPEG, or WebP format. Its documented strengths include practical image creation, inpainting, transparent-background assets, and control over image quality and dimensions. GPT-Image-1 is deprecated and scheduled to shut down on October 23, 2026.

What is GPT-Image-1?

GPT-Image-1 is OpenAI’s image-generation and image-editing model for API use. It accepts text prompts and image inputs, then produces image outputs. In practical terms, a developer can use it to create an illustration from a description, revise an existing image, preserve visual elements from reference images, or replace a selected area through a mask-based editing workflow.

OpenAI released GPT-Image-1 on April 23, 2025. The company described it as the API model behind the native image-generation experience introduced in ChatGPT. Its intended role is visual content production rather than general conversation, reasoning, coding, speech, video generation, or embeddings.

The model is currently listed as deprecated. OpenAI’s model documentation states that API access is scheduled to end on October 23, 2026. That lifecycle status is one of the most important practical considerations: GPT-Image-1 can still be relevant for understanding or maintaining an existing integration, but it is a less suitable foundation for a new long-lived product unless a migration path has been considered.

Core image-generation and editing capabilities

GPT-Image-1 supports both text-to-image generation and image editing. A text-only request can describe the desired subject, composition, style, or design. When an image is supplied, it can act as a reference for the generated result or as the source for an edit. Masks allow an application to identify the portion of an image that should be changed, which is useful for inpainting-style workflows such as replacing an object, altering a background, or correcting a localized area.

  • Text-to-image generation: Create a new image from a written prompt.
  • Reference-image workflows: Use supplied images to guide appearance, composition, or identity.
  • Mask-based editing: Target particular regions instead of changing the entire image.
  • Transparent backgrounds: Generate assets with transparency when a transparent background is requested and PNG or WebP output is selected.
  • Quality controls: Select low, medium, high, or automatic quality.
  • Output formats: Receive PNG, JPEG, or WebP images, with optional JPEG and WebP compression.

These features make the model more useful for production workflows than a generator that only creates a new image from a prompt. For example, an e-commerce application could create product artwork, revise a supplied product image, or generate a transparent asset for placement in a catalog.

Supported sizes, quality, and formats

The documented image sizes are 1024×1024, 1024×1536, and 1536×1024. An automatic size option is also available. The square setting is suitable for many social, thumbnail, and product-card uses, while the portrait and landscape settings provide more room for posters, banners, and editorial compositions.

Quality can be set to low, medium, high, or auto. Higher quality generally carries a higher image-generation charge, so the setting should match the task. Low quality can be appropriate for drafts or rapid iteration, whereas high quality may be justified for a final marketing asset. The provider documents the following per-image generation prices:

Quality1024×10241024×1536 or 1536×1024
Low$0.011 per image$0.016 per image
Medium$0.042 per image$0.063 per image
High$0.167 per image$0.25 per image

PNG, JPEG, and WebP are supported. JPEG and WebP compression can be configured where appropriate, while transparent backgrounds require an explicit transparent-background request together with PNG or WebP output.

GPT-Image-1 pricing and cost planning

GPT-Image-1 pricing has more than one component. The per-image prices above cover the image-generation charge, but a request can also incur charges for text tokens and image tokens. OpenAI documents these rates:

  • Text input: $5 per million tokens.
  • Cached text input: $1.25 per million tokens.
  • Image input: $10 per million image tokens.
  • Cached image input: $2.50 per million image tokens.
  • Image output: $40 per million image tokens.

This distinction matters most for editing and reference-image workflows. A request that supplies one or more images can incur image-input charges in addition to the selected generation price. Applications that repeatedly reuse the same inputs may benefit from the documented cached-input rates, but the final cost still depends on the request’s token usage and image settings.

From a cost perspective, low-quality square images are the least expensive documented generation option. High-quality portrait or landscape images are the most expensive per-image option. These are provider-published prices, not an estimate of total application cost; storage, retries, orchestration, and any other application expenses are separate.

Inputs, outputs, and unsupported features

GPT-Image-1 has multimodal input because it can process both text and images. Its direct output is image data, not text, audio, or video. The supplied model documentation does not publish a context-length limit or maximum-output-token limit for this model; those values should therefore be treated as undocumented rather than assumed to be unlimited.

GPT-Image-1 is not documented as a general-purpose language model. It does not support text generation as its primary output, and it should not be selected for ordinary writing, coding, long-form reasoning, speech, video, or embedding workloads. The model documentation also lists the following unsupported features:

  • Streaming
  • Function calling
  • Structured outputs
  • Fine-tuning

There is consequently no documented reasoning score or coding capability to evaluate in the same way as a text-generation model. Any reasoning involved in interpreting an image prompt should not be confused with a supported general reasoning interface.

Strengths and limitations

Where GPT-Image-1 is strongest

GPT-Image-1’s main strength is the combination of generation and editing in one API-oriented image model. It can create new visual assets, use reference images, perform localized mask-based changes, and return common web-friendly formats. The documented quality and size controls also let developers trade visual quality, dimensions, and cost for different stages of a workflow.

It is particularly well suited to marketing graphics, e-commerce imagery, illustrations, concept art, educational visuals, gaming assets, and applications that let users revise their own images. Transparent-background support is useful for logos, product cutouts, icons, characters, and other assets that need to be placed on different backgrounds.

Important limitations

The largest limitation is its lifecycle. GPT-Image-1 is deprecated and scheduled to shut down on October 23, 2026. A team choosing it for a new application should plan for replacement rather than treating the model as a permanent endpoint.

It also lacks several capabilities that may matter in a broader automation system. There is no streaming, function calling, structured-output interface, or fine-tuning support documented for this exact model. It is not a text, speech, video, or embedding model, so a product that needs those functions will require additional models or services. Editing and reference-image requests can also cost more than the headline per-image price because image tokens are billed separately.

When to choose GPT-Image-1

GPT-Image-1 is a reasonable choice when the central requirement is API-based image creation or editing and the project can accommodate its announced shutdown date. It is especially appropriate when the workflow needs one or more of the following:

  • Generating images from natural-language descriptions.
  • Editing user-provided images with references or masks.
  • Creating marketing, product, educational, or game-related visuals.
  • Producing transparent-background PNG or WebP assets.
  • Choosing between low, medium, and high generation quality.
  • Using square, portrait, or landscape output dimensions.

It is less appropriate when the application needs a general conversational model, coding or reasoning, real-time streaming, function calls, structured JSON responses, audio or video, or a stable long-term model without migration work. In those cases, a current model designed for the required modality or tool interface is more suitable. The supplied research does not identify a specific successor by name, so comparisons should be made against the current GPT Image offering documented by OpenAI rather than assuming an unverified replacement.

Practical evaluation before adoption

Before integrating GPT-Image-1, test the exact prompts and image-editing cases that matter to the product. Compare low, medium, and high quality at the required dimensions, and measure the resulting per-request cost rather than relying only on the generation fee. Include reference-image and mask-based requests in testing because their image-token usage can change the economics.

Also verify that the application can handle image responses in the selected format, transparent-background requirements, compression settings, and the absence of streaming or structured outputs. Most importantly, document a migration plan before launch. Because OpenAI has announced a shutdown date, a successful short-term integration still needs an exit strategy.


Answers to Frequently Asked Questions

What features does GPT-Image-1 not support?
GPT-Image-1 is designed for image generation and editing rather than general-purpose AI tasks. It does not support streaming, function calling, structured outputs, or fine-tuning, and it is not intended for text, coding, reasoning, speech, video, or embedding workloads.
What image sizes, formats, and quality levels does GPT-Image-1 support?
GPT-Image-1 supports 1024×1024, 1024×1536, 1536×1024, and automatic sizing. Quality can be set to low, medium, high, or auto. Supported output formats are PNG, JPEG, and WebP; transparent backgrounds require PNG or WebP.
How much does GPT-Image-1 cost?
Image-generation prices range from $0.011 to $0.25 per image, depending on quality, size, and orientation. Requests may also incur separate text-token, image-input-token, and image-output-token charges, especially when reference images or editing workflows are used.
What is GPT-Image-1 used for?
GPT-Image-1 is OpenAI’s API model for generating and editing images. It can create images from text prompts, use reference images, apply mask-based edits, and produce assets with transparent backgrounds.
Is GPT-Image-1 still available, and when will it shut down?
GPT-Image-1 is currently listed as deprecated. OpenAI’s documentation states that API access is scheduled to end on October 23, 2026, so new integrations should include a migration plan.


Sources 5
Provider

About OpenAI