Chat Latest

Chat Latest

by OpenAI · Current rolling alias; underlying model snapshot is regularly updated

Chat Latest is OpenAI’s rolling alias for the latest Instant model used in ChatGPT. It accepts text and image input, returns text, supports a 400,000-token context window and 128,000-token maximum output, and offers streaming, function calling, structured outputs, web search, file search, code interpreter, MCP, and batch processing. Its regularly changing snapshot makes it convenient for current ChatGPT-style work but less suitable for version-sensitive production deployments.

Text Reasoning Coding
Chat Latest is OpenAI’s current alias for the latest Instant model used in ChatGPT. It is intended to provide convenient access to the current ChatGPT Instant experience, including text conversations, image-aware interactions, and tool-assisted workflows. The important qualification is that Chat Latest is a moving alias: OpenAI regularly updates the model snapshot behind the name. That makes it useful when staying current matters more than version stability, but it is a weaker choice for production systems that need tightly controlled behavior over time.
Outputs

What Chat Latest can produce

Text
Inputs

What it can understand

Text Images Multimodal input
Capabilities

Supported features

Tool use Web search Streaming Structured output Prompt caching Batch API
Model profile

Performance characteristics

8/10 Reasoning
8/10 Coding
9/10 Speed
6/10 Cost efficiency
Specifications

Technical details

Model family Chat Latest
Model type General Purpose
Context window 400K tokens
Maximum output 128K tokens
Knowledge cutoff August 31, 2025
Status Current rolling alias; underlying model snapshot is regularly updated
Knowledge cutoff notes

The documented knowledge cutoff applies to the current Chat Latest model page, while the underlying snapshot may change regularly because chat-latest is a rolling alias.

Model notes

Chat Latest is the chat-latest alias for the latest Instant model currently used in ChatGPT. OpenAI says the underlying snapshot will be regularly updated. The model page recommends GPT-6 Astra for production API usage. Pricing is listed per 1 million tokens, with separate cached-input pricing. Structured outputs are supported, but the documentation does not separately confirm a legacy JSON-mode capability.

Cost

Model pricing

Input $5.00 per 1 million input tokens; cached input $0.50 per 1 million tokens
Output $30.00 per 1 million output tokens
Model guide

Chat Latest: OpenAI’s Rolling Instant Model for ChatGPT

Chat Latest is OpenAI’s rolling alias for the latest Instant model used in ChatGPT. It combines fast text generation with image input, a 400,000-token context window, tool use, structured outputs, and support for workflows such as web search, file search, code interpreter, image generation tools, MCP, and batch processing. Its defining trade-off is that the underlying snapshot changes regularly, making it convenient for current ChatGPT-style work but less suitable for applications that require stable, reproducible production behavior.

What is Chat Latest?

Chat Latest is an OpenAI model alias that points to the latest Instant model currently used in ChatGPT. Unlike a versioned model identifier, the name does not permanently identify one fixed set of model weights or behavior. OpenAI states that the underlying snapshot is updated regularly, so the model may change while the chat-latest alias remains the same.

Its primary purpose is to provide a current ChatGPT-style model for everyday conversations and general-purpose work. Typical tasks include writing, summarization, analysis, image-aware discussion, and research assisted by OpenAI tools. The alias is therefore more convenient than version-specific when the priority is automatically receiving the latest Instant snapshot.

Chat Latest sits in a different practical category from a version-locked production target. The supplied OpenAI model guidance recommends GPT-6 Astra for production API usage instead. That does not make Chat Latest unsuitable for every API workflow, but it does mean developers should treat it as a changing target rather than a long-term compatibility contract.

Input, output, and context limits

Chat Latest accepts text and image input and produces text output. In practical terms, you can provide written instructions or an image for the model to analyze, but the model does not itself return native images, audio, or video. Image-generation tools may be available in supported workflows, but using a tool is different from the underlying model having direct image output.

SpecificationDocumented value
Text inputSupported
Image inputSupported
Audio inputNot supported
Video inputNot supported
Text outputSupported
Native image, audio, or video outputNot supported
Context window400,000 tokens
Maximum output128,000 tokens

A token is a unit of text used by the model; it may represent a whole word, part of a word, punctuation, or other text. The 400,000-token context window is large enough for substantial instructions, documents, and conversation history, subject to the limits of the specific application sending the request. The 128,000-token maximum output is a separate limit: it describes how much text the model can generate in one response, not how much information it can receive.

Tools, structured responses, and API features

Chat Latest supports streaming, which allows an application to display generated text progressively instead of waiting for the complete response. It also supports function calling, a mechanism that lets the model request an application-defined function with structured arguments. The application remains responsible for executing the function and returning its result.

The documented tool support includes web search, file search, image-generation tools, code interpreter, and MCP. These capabilities can extend a conversation beyond text generation. For example, web search can support research workflows, file search can retrieve information from supplied files, and code interpreter can assist with executable analysis in supported environments. MCP support can connect the model to compatible external tools or services. Availability and behavior still depend on the API surface and configuration being used.

Structured outputs are supported, allowing responses to be constrained to a specified structure where the relevant OpenAI interface supports it. The supplied documentation does not separately confirm a legacy JSON-mode capability, so structured outputs should not automatically be treated as proof of every distinct JSON feature.

OpenAI lists Chat Completions, Responses, Batch, and other endpoint categories for the model. Batch processing is supported. Fine-tuning is not supported, so organizations cannot use the documented model as a fine-tuning base for creating a customized version of Chat Latest.

Pricing for Chat Latest

OpenAI lists the following usage prices:

  • Input: $5.00 per 1 million tokens
  • Cached input: $0.50 per 1 million tokens
  • Output: $30.00 per 1 million tokens

Cached-input pricing applies to eligible cached input rather than replacing the standard input rate in every request. Output is priced substantially higher per token than ordinary input, so applications that generate long responses should account for both response length and the number of requests they make. Batch API access is listed as supported, which may be useful for workloads that can be processed asynchronously rather than returned immediately.

The price figures above are provider-listed token rates, not a monthly subscription price. Actual usage cost depends on input volume, output volume, caching eligibility, and the tools or services involved in a complete workflow.

Main strengths and trade-offs

Chat Latest’s clearest strength is its combination of current ChatGPT positioning, fast response expectations, broad tool support, and a large context window. It can handle ordinary text work, inspect images, stream responses, call functions, and participate in research or analysis workflows without requiring the user to switch to a different model alias whenever OpenAI updates the Instant snapshot.

The supplied editorial evaluation gives Chat Latest a reasoning score of 8 out of 10, a coding score of 8 out of 10, a speed score of 9 out of 10, and a cost score of 6 out of 10. These are editorial assessments, not OpenAI-published benchmark results or official provider ratings. They suggest a profile oriented toward fast general-purpose assistance with solid reasoning and coding utility, while its token pricing is less attractive than that of lower-cost models for very high-volume work.

The principal trade-off is stability. Because the snapshot changes regularly, the same prompt may not always produce identical behavior, formatting, or quality over time. This matters for regression tests, benchmark tracking, carefully tuned prompts, compliance-sensitive workflows, and applications whose users depend on stable outputs. A versioned model identifier is more appropriate when reproducibility is more important than automatically receiving the latest Instant model.

Best uses for Chat Latest

  • ChatGPT-style conversations: It is suited to everyday questions, drafting, rewriting, summarization, and general knowledge work.
  • Image-aware assistance: Users can combine written instructions with image input for analysis and discussion.
  • Tool-assisted research: Web search and file search can support workflows that need information beyond the initial prompt.
  • Rapid application prototypes: Developers can test streaming, function calling, structured outputs, and supported tools without selecting a fixed Instant snapshot first.
  • Long-context work: The documented 400,000-token context window can accommodate large instructions, document collections, or extended conversations when the surrounding application supports them.
  • Asynchronous bulk processing: Batch support can fit jobs that do not require an immediate response.

When should you choose Chat Latest?

Choose Chat Latest when you want the current ChatGPT Instant experience and value speed, multimodal input, and integrated tools more than a frozen model version. It is a reasonable fit for interactive assistants, writing and analysis products, image-aware text applications, research tools, and prototypes that should track OpenAI’s latest Instant snapshot automatically.

It is less appropriate when your application needs predictable behavior across deployments or when a prompt and output format must remain stable for months. In those cases, select a versioned model instead of a rolling alias. The supplied guidance specifically points to GPT-6 Astra for production API usage, so teams building a durable production API service should evaluate that recommendation rather than assuming Chat Latest is the default choice.

Another model type may also be preferable when the main priority is minimizing cost. Chat Latest’s listed output rate of $30 per 1 million tokens can make extensive generation expensive compared with lower-cost options, although the correct alternative depends on the quality, context, tool, and latency requirements of the workload. Conversely, a specialized model may be more suitable if the task requires audio or video input, native non-text output, or fine-tuning, because those capabilities are not supported by Chat Latest as documented.

Limitations to consider

Chat Latest has four important limitations. First, it is a rolling alias, so its behavior can change without a name change. Second, it returns text rather than native image, audio, or video output, even though supported tools can broaden what an application accomplishes. Third, it does not support fine-tuning. Fourth, its pricing, especially for output, may be difficult to justify for large-scale generation when a less expensive model can meet the quality and capability requirements.

These limitations do not prevent useful deployments, but they affect model selection. Use the alias for current, flexible, interactive work; use a versioned production model when repeatability and controlled change are essential. Before committing to an implementation, test the specific prompting, tool calls, structured response requirements, and cost profile that matter to your application.


Answers to Frequently Asked Questions

When should developers use Chat Latest instead of a versioned model?
Use Chat Latest when you want the current ChatGPT Instant experience, fast responses, image input, and integrated tools. Choose a versioned model when stable behavior, reproducible outputs, long-term compatibility, or controlled changes are more important. The supplied guidance recommends GPT-6 Astra for production API usage.
What tools and API features does Chat Latest support?
Chat Latest supports streaming, function calling, structured outputs, Batch processing, and endpoints including Chat Completions and Responses. Documented tools include web search, file search, image-generation tools, code interpreter, and MCP, depending on the API surface and configuration.
What are Chat Latest’s context window and pricing limits?
Chat Latest has a documented context window of 400,000 tokens and a maximum output of 128,000 tokens. OpenAI lists pricing of $5 per 1 million input tokens, $0.50 per 1 million cached input tokens, and $30 per 1 million output tokens.
What is Chat Latest?
Chat Latest is an OpenAI model alias that points to the latest Instant model used in ChatGPT. Its underlying model snapshot is updated regularly, so the alias does not represent one permanently fixed set of model weights or behavior.
What inputs and outputs does Chat Latest support?
Chat Latest supports text and image input and produces text output. It does not natively support audio or video input, or native image, audio, or video output, although tools may extend its capabilities in supported workflows.


Sources 4
Provider

About OpenAI