GPT-5.1

GPT-5.1 Chat

by OpenAI · Retired; the API alias gpt-5.1-chat-latest was shut down on 2026-07-23. GPT-5.1 models were retired from ChatGPT on 2026-03-11.

GPT-5.1 Chat was OpenAI’s ChatGPT-optimized API alias for the GPT-5.1 snapshot. It accepted text and images, generated text, supported function calling, structured outputs, streaming, and batch processing, and offered a 128,000-token context window. It was retired from ChatGPT on March 11, 2026, and its API alias shut down on July 23, 2026.

Text Reasoning Coding
GPT-5.1 Chat was the API identity for the GPT-5.1 snapshot used in ChatGPT. It was designed for conversational applications that needed natural instruction following, image understanding, tool use, and relatively fast text responses. The model accepted text and image input but produced text only. Its historical API pricing was $1.25 per million input tokens and $10 per million output tokens, with discounted cached input. GPT-5.1 Chat is no longer suitable for new deployments because OpenAI retired the ChatGPT model and later shut down the API alias.
Outputs

What GPT-5.1 Chat can produce

Text
Inputs

What it can understand

Text Images Multimodal input
Capabilities

Supported features

Tool use Streaming Structured output Prompt caching Batch API
Model profile

Performance characteristics

8/10 Reasoning
8/10 Coding
8/10 Speed
7/10 Cost efficiency
Specifications

Technical details

Model family GPT-5.1
Model type General Purpose
Context window 128K tokens
Maximum output 16K tokens
Knowledge cutoff 2024-09-30
Release date 2025-11-13
Status Retired; the API alias gpt-5.1-chat-latest was shut down on 2026-07-23. GPT-5.1 models were retired from ChatGPT on 2026-03-11.
Deprecation date 2026-04-22
Shutdown date 2026-07-23
Knowledge cutoff notes

OpenAI documented September 30, 2024 as the model's knowledge cutoff. External tools, retrieval, or user-provided context could supply newer information but did not change the underlying cutoff.

Model notes

GPT-5.1 Chat was the API alias for the GPT-5.1 snapshot used in ChatGPT. Its canonical API identifier was gpt-5.1-chat-latest. OpenAI documented text and image input, text output, a 128,000-token context window, and a 16,384-token maximum output. It supported streaming, function calling, structured outputs, and batch processing, but not fine-tuning. The documented knowledge cutoff was September 30, 2024. The model was retired in ChatGPT on March 11, 2026 and its API access ended on July 23, 2026. Editorial scores are comparative estimates rather than vendor-published ratings.

Cost

Model pricing

Input $1.25 per 1 million input tokens; cached input $0.125 per 1 million tokens
Output $10.00 per 1 million output tokens
Model guide

GPT-5.1 Chat: OpenAI’s Retired ChatGPT-Optimized API Model

GPT-5.1 Chat was OpenAI’s ChatGPT-optimized API model, exposed as gpt-5.1-chat-latest. It accepted text and images, generated text, supported streaming, function calling, structured outputs, and batch processing, and provided a 128,000-token context window with a 16,384-token maximum output. It was retired from ChatGPT on March 11, 2026, and its API alias was shut down on July 23, 2026.

What GPT-5.1 Chat was

GPT-5.1 Chat was OpenAI’s ChatGPT-optimized API model. Developers accessed it through the moving alias gpt-5.1-chat-latest, which represented the GPT-5.1 snapshot used in ChatGPT rather than a completely separate generation family. OpenAI released GPT-5.1 in November 2025, with GPT-5.1 Chat serving the conversational role within that family.

The model was intended for applications where the quality of dialogue, instruction following, and response speed mattered more than using a specialized image, audio, or video generation system. Typical examples included customer-support assistants, general chat interfaces, visual question answering, document interpretation, and tool-using workflows.

There is an important access distinction: GPT-5.1 Chat’s API identity was related to ChatGPT but was not the same thing as a ChatGPT subscription plan. The model was available through OpenAI’s Chat Completions and Responses APIs, and batch processing was also supported.

Status and place in OpenAI’s catalog

GPT-5.1 Chat is a retired model rather than an active choice for a new production integration. OpenAI retired GPT-5.1 models from ChatGPT on March 11, 2026. The gpt-5.1-chat-latest API alias remained available after that ChatGPT retirement but was scheduled for API shutdown on July 23, 2026. OpenAI listed GPT-5.6 Sol as the recommended replacement in the supplied model information.

Because gpt-5.1-chat-latest was a moving alias, it was less appropriate for workflows that required a permanently reproducible model target. Where a dated snapshot was available and supported, using that snapshot would have provided more stable behavior. After the shutdown date, applications need to migrate to an active model rather than continue targeting this alias.

Input and output modalities

GPT-5.1 Chat accepted both text and images as input and returned text as output. Image understanding enabled tasks such as asking questions about a photograph, extracting meaning from a document image, examining a screenshot, or grounding a conversation in visual information.

It was not an audio or video model. The documented capabilities did not include audio input, video input, or native image, audio, video, music, speech, or embedding output. In practical terms, an application could send an image alongside a text prompt, but the model’s response was still text. A separate service would have been needed for media generation or speech output.

Technical limits and supported capabilities

SpecificationDocumented value
ProviderOpenAI
API identifiergpt-5.1-chat-latest
Model familyGPT-5.1
Context window128,000 tokens
Maximum output16,384 tokens
InputText and images
OutputText
StreamingSupported
Function callingSupported
Structured outputsSupported
Fine-tuningNot supported
Knowledge cutoffSeptember 30, 2024

The 128,000-token context window determined how much combined prompt, conversation history, image-related input, and other supported context could be supplied in one request. The 16,384-token output limit restricted the length of one generated response; it did not mean every response would approach that size.

Streaming allowed an application to display generated text incrementally instead of waiting for the complete response. Function calling allowed the model to request an application-defined operation, such as looking up an account or querying a database, while the application remained responsible for executing that operation and returning the result.

Structured outputs were useful when a response needed to follow a supplied schema, for example when extracting fields from a document or returning consistent records to another program. The supplied research confirms structured outputs but does not verify a separate legacy JSON-mode capability, so those features should not be treated as interchangeable.

Historical pricing and API access

OpenAI listed GPT-5.1 Chat at $1.25 per one million input tokens and $10 per one million output tokens. Cached input was listed at $0.125 per one million tokens. These were usage prices for the API, not recurring ChatGPT subscription prices.

The input and output rates created a meaningful cost difference between short conversational requests and workflows that generated long answers at scale. Reusing eligible prompt content through cached input could reduce the input portion of usage costs, but output tokens remained charged at the documented output rate. Since the model and alias were retired, these prices should be understood as historical pricing rather than a current purchasing option.

Reasoning, coding, speed, and cost trade-offs

GPT-5.1 Chat was positioned primarily as a conversational model rather than a dedicated reasoning or coding model. Its practical strengths were instruction following, natural dialogue, image-grounded conversation, streaming, and integration with external tools. It could assist with coding questions and software workflows, especially where the task involved explaining code, extracting structured information, or coordinating an API action, but the supplied research does not identify it as a specialized coding model.

Editorial scoring in the supplied data rates reasoning, coding, and speed at 8 out of 10, with cost rated at 7 out of 10. These are comparative editorial estimates, not OpenAI-published benchmark results or guarantees. They suggest a model viewed as balanced for general-purpose use, but they should not be presented as formal performance measurements.

Compared with a slower, more deliberate reasoning configuration, GPT-5.1 Chat was the more natural fit when responsive conversation and broad task coverage mattered. Compared with a smaller or cheaper model, its historical pricing could be harder to justify for high-volume, simple classification or short-answer workloads. Those trade-offs are useful for understanding its original positioning, but retirement now outweighs them for new deployments.

Best use cases

  • Conversational assistants: Build chat experiences that need multi-turn instruction following and natural text responses.
  • Customer-support workflows: Combine the model with function calls for account lookups, ticket operations, or knowledge retrieval.
  • Image-grounded chat: Analyze screenshots, photographs, or document images while explaining the result in text.
  • Structured extraction: Return fields that conform to a defined schema when processing text or visual documents.
  • Streaming interfaces: Show partial responses as they are generated for a more responsive user experience.
  • Tool-using applications: Let the model choose among application-provided functions while the surrounding software controls execution.
  • Batch processing: Process supported workloads asynchronously where immediate interactive responses were not required.

When GPT-5.1 Chat would have been the right choice

Before retirement, GPT-5.1 Chat made sense when an application needed a single text-generation model that combined image input, conversational behavior, function calling, streaming, and structured responses. It was especially suitable when a product needed visual understanding but did not need the model itself to generate media.

A different type of option would have been more appropriate for native image, audio, or video generation; for audio or video understanding; or for fine-tuning on organization-specific examples. A lower-cost model could have been preferable for large volumes of routine requests, while a more specialized reasoning or coding option could have been preferable for tasks where those capabilities were the primary requirement. For a new project today, however, the decisive alternative is an active replacement, including the recommended model identified by OpenAI, rather than GPT-5.1 Chat itself.

Limitations and caveats

  • GPT-5.1 Chat generated text only and did not natively produce images, audio, video, music, speech, or embeddings.
  • It did not support audio or video input.
  • Fine-tuning was not available for this model.
  • Its documented knowledge cutoff was September 30, 2024. External retrieval or application-provided context could supply newer information, but would not change the underlying cutoff.
  • The moving API alias could change behavior over time, making it less suitable for strict reproducibility than a supported dated snapshot.
  • The model was retired from ChatGPT on March 11, 2026 and the API alias was shut down on July 23, 2026.

In summary, GPT-5.1 Chat was a balanced, chat-focused GPT-5.1 API model with image understanding, tool support, structured outputs, and a large context window. Those characteristics explain its original usefulness, but its retirement means the model is now primarily a historical reference point rather than a deployable OpenAI option.


Answers to Frequently Asked Questions

How much did GPT-5.1 Chat cost through the API?
The historical API pricing was $1.25 per one million input tokens, $10 per one million output tokens, and $0.125 per one million cached input tokens. These prices are no longer a current purchasing option because the model and API alias were retired.
What were the technical limits of GPT-5.1 Chat?
GPT-5.1 Chat had a 128,000-token context window and a maximum output of 16,384 tokens. It supported streaming, function calling, structured outputs, and batch processing, but did not support fine-tuning. Its documented knowledge cutoff was September 30, 2024.
What inputs and outputs did GPT-5.1 Chat support?
GPT-5.1 Chat accepted text and images as input and generated text as output. It could analyze photographs, screenshots, and document images, but it did not support audio or video input or native image, audio, video, music, speech, or embedding output.
Is GPT-5.1 Chat still available?
No. GPT-5.1 Chat was retired from ChatGPT on March 11, 2026, and the gpt-5.1-chat-latest API alias was scheduled for shutdown on July 23, 2026. OpenAI listed GPT-5.6 Sol as the recommended replacement.
What was GPT-5.1 Chat?
GPT-5.1 Chat was OpenAI’s ChatGPT-optimized API model in the GPT-5.1 family. Developers accessed it through the moving alias gpt-5.1-chat-latest for conversational applications requiring instruction following, image understanding, streaming, function calling, and text generation.


Sources 6
Provider

About OpenAI