Model catalog

OpenAI Models

Browse the AI models associated with OpenAI. Compare current and historical models by family, capabilities, context window, availability and intended use.

107 models tracked
107 Total models
50 Model families
13 Model types
63 Current / accessible
All models

OpenAI model catalog

OpenAI logo
Chat Latest

Chat Latest

ChatGPT-style instant responses, general-purpose writing and analysis, image-aware conversations, and tool-assisted workflows

Type General Purpose
Context 400K
Reasoning 8/10
Speed 9/10
Multimodal Image input Tool use
Status

Current rolling alias; underlying model snapshot is regularly updated

Input $5.00 per 1 million input tokens; cached input $0.50 per 1 million tokens
Output $30.00 per 1 million output tokens
View model →
OpenAI logo
CLIP

CLIP

Zero-shot image classification, image-text similarity, semantic image retrieval, multimodal indexing, and computer-vision research

Type Multimodal
Context 77
Reasoning 2/10
Speed 7/10
Multimodal Image input
Status

Public research release with downloadable weights; not verified as a current OpenAI hosted API model

View model →

Codex CLI coding workflows, code question answering, code editing, repository tasks, and low-latency software-engineering assistance

Type Coding
Context 200K
Reasoning 7/10
Speed 8/10
Multimodal Image input Tool use
Status

Retired; API access ended on 2026-02-12

Input $1.50 per 1M input tokens; $0.375 per 1M cached input tokens
Output $6.00 per 1M output tokens
View model →
OpenAI logo
Computer-Using Agent

computer-use-preview

Controlled browser automation, computer-use research, UI testing, and repetitive interface workflows

Type Other
Context 8K
Reasoning 6/10
Speed 7/10
Multimodal Image input Tool use
Status

Deprecated

Input $3.00 per 1 million input tokens; tool-specific computer-use calls may incur separate fees
Output $12.00 per 1 million output tokens
View model →
OpenAI logo
DALL·E

DALL·E 2

Historical research on text-to-image generation, legacy image workflows, and comparisons with newer OpenAI image models.

Type Other
Multimodal Image input Media output
Status

Retired; deprecated and removed from the OpenAI API on May 12, 2026.

View model →
OpenAI logo
DALL·E

DALL·E 3

Historical text-to-image generation, concept art, illustration, visual ideation, marketing imagery, and prompt-following research

Type Other
Reasoning 1/10
Speed 6/10
Media output
Status

Retired; deprecated and removed from the OpenAI API on May 12, 2026

Input Not applicable to current use; historical pricing was charged per generated image rather than per input token
Output Historical API pricing started at $0.04 per 1024×1024 standard-quality image; higher prices applied to HD and larger formats
View model →

Maintaining legacy text-completion applications, historical GPT-3 base-model behavior, and existing compatible fine-tuned workflows before shutdown

Type Lightweight
Reasoning 2/10
Speed 7/10
Status

Deprecated; currently accessible but scheduled to shut down on 2026-09-28

Input $0.40 per 1 million input tokens
Output $0.40 per 1 million output tokens
View model →

Legacy text completion, code continuation, and inference from existing davinci-002 fine-tuned models before shutdown

Type General Purpose
Reasoning 3/10
Speed 6/10
Status

Deprecated; API access scheduled to shut down on September 28, 2026

Input $2.00 per 1M tokens
Output $2.00 per 1M tokens
View model →

Low-cost, high-volume text generation, summarization, classification, extraction, simple chatbots, and legacy API integrations

Type General Purpose
Context 16K
Reasoning 4/10
Speed 8/10
Fine-tuning
Status

Deprecated; still available through the OpenAI API

Input $0.50 per 1 million input tokens
Output $1.50 per 1 million output tokens
View model →
OpenAI logo
GPT-4

GPT-4

Maintaining established GPT-4 integrations, general-purpose text generation, analysis, writing, and coding workloads

Type General Purpose
Context 8K
Reasoning 8/10
Speed 5/10
Multimodal Image input Tool use
Status

Legacy; older high-intelligence GPT model

Input $30 per 1 million prompt tokens
Output $60 per 1 million completion tokens
View model →

Legacy high-context text and image analysis, function calling, JSON-mode workflows, and existing GPT-4 Turbo integrations

Type Multimodal
Context 128K
Reasoning 7/10
Speed 6/10
Multimodal Image input Tool use
Status

Deprecated but currently accessible; scheduled for shutdown on October 23, 2026

Input $10 per 1 million input tokens
Output $30 per 1 million output tokens
View model →

Historical long-context text generation, document analysis, structured text generation, and general-purpose assistant applications.

Type General Purpose
Context 128K
Reasoning 7/10
Speed 7/10
Fine-tuning
Status

Retired; the gpt-4-turbo-preview alias pointed to gpt-4-0125-preview, which was shut down on 2026-03-26.

Input $10 per 1 million tokens
Output $30 per 1 million tokens
View model →
OpenAI logo
GPT-4.1

GPT-4.1

Software engineering, long-context document analysis, precise instruction following, structured extraction, tool-enabled agents, and image understanding

Type General Purpose
Context 1.05M
Reasoning 7/10
Speed 8/10
Multimodal Image input Tool use
Status

Current; default GPT-4.1 alias with gpt-4.1-2025-04-14 snapshot

Input $2.00 per 1M input tokens; $0.50 per 1M cached input tokens
Output $8.00 per 1M output tokens
View model →

Fast, cost-efficient instruction following, coding assistance, image understanding, structured extraction, tool calling, and long-context API applications

Type Lightweight
Context 1.05M
Reasoning 6/10
Speed 9/10
Multimodal Image input Tool use
Status

current

Input $0.40 per 1 million tokens; cached input $0.10 per 1 million tokens
Output $1.60 per 1 million tokens
View model →

High-volume, latency-sensitive classification, extraction, routing, summarization, lightweight assistants, image-assisted analysis, and simple tool-calling workflows

Type Lightweight
Context 1.05M
Reasoning 4/10
Speed 10/10
Multimodal Image input Tool use
Status

Deprecated; currently accessible as of September 23, 2026; scheduled for shutdown on October 23, 2026

Input $0.10 per 1M input tokens; $0.025 per 1M cached input tokens
Output $0.40 per 1M output tokens
View model →

Historical general-purpose writing, creative work, nuanced communication, image understanding, programming assistance, and applications needing function calling or structured outputs.

Type General Purpose
Context 128K
Reasoning 7/10
Speed 3/10
Multimodal Image input Tool use
Status

Retired from the API on 2025-07-14; retired from ChatGPT in June 2026

Input $75.00 per 1 million input tokens; $37.50 per 1 million cached input tokens
Output $150.00 per 1 million output tokens
View model →

Fast general-purpose conversations, vision, voice interactions, coding, and everyday productivity

Type Multimodal
Context 128K
Reasoning 8/10
Speed 9/10
Multimodal Image input Audio input
Status

Active

View model →
OpenAI logo
GPT-4o

GPT-4o

General-purpose assistants, image understanding, coding help, structured extraction, multilingual generation, and latency-sensitive API workflows

Type Multimodal
Context 128K
Reasoning 8/10
Speed 9/10
Multimodal Image input Tool use
Status

Current in the OpenAI API; retired from ChatGPT on 2026-02-13. The gpt-4o-2024-05-13 snapshot is scheduled for API shutdown on 2026-10-23.

Input $2.50 per 1M input tokens; $1.25 per 1M cached input tokens
Output $10.00 per 1M output tokens
View model →

Voice assistants, spoken conversational agents, audio-enabled customer service, and applications requiring direct audio understanding and speech generation

Type Multimodal
Context 128K
Reasoning 7/10
Speed 6/10
Multimodal Audio input Media output
Status

Retired; API access ended May 7, 2026

Input $2.50 per 1M text input tokens; $40 per 1M audio input tokens
Output $10.00 per 1M text output tokens; $80 per 1M audio output tokens
View model →

Low-cost, high-volume text and image understanding, classification, extraction, translation, tagging, customer support, routing, and structured data generation

Type Lightweight
Context 128K
Reasoning 5/10
Speed 9/10
Multimodal Image input Tool use
Status

Current canonical model alias; dated snapshot gpt-4o-mini-2024-07-18 is available

Input $0.15 per 1M input tokens
Output $0.60 per 1M output tokens
View model →

Lower-cost audio understanding, conversational voice interfaces, and applications requiring text and spoken-audio input/output

Type Multimodal
Context 128K
Reasoning 3/10
Speed 8/10
Multimodal Audio input Media output
Status

Deprecated; scheduled for API shutdown on 2027-01-20

Input Text: $0.15 per 1M tokens; audio: $10.00 per 1M tokens
Output Text: $0.60 per 1M tokens; audio: $20.00 per 1M tokens
View model →

Low-cost realtime voice assistants, speech-to-speech interfaces, interactive audio applications, and conversational prototypes

Type Lightweight
Context 16K
Reasoning 4/10
Speed 9/10
Multimodal Audio input Media output
Status

Deprecated; scheduled for API removal on 2027-01-20

Input Text: $0.60 per 1M input tokens; audio: $10.00 per 1M audio tokens; cached input: $0.30 per 1M tokens
Output Text: $2.40 per 1M output tokens; audio: $20.00 per 1M audio tokens
View model →

Low-latency voice assistants, speech-to-speech applications, live translation, language learning, and interactive customer support

Type Multimodal
Context 32K
Reasoning 6/10
Speed 9/10
Multimodal Audio input Media output
Status

Retired; API access ended 2026-05-07

Input $5 per 1M text tokens; $40 per 1M audio tokens; cached text and audio input $2.50 per 1M tokens
Output $20 per 1M text tokens; $80 per 1M audio tokens
View model →

Historical web-search applications built around OpenAI Chat Completions

Type Other
Context 128K
Reasoning 5/10
Speed 6/10
Tool use Streaming
Status

Retired; shut down on 2026-07-23

Input $2.50 per 1 million input tokens; historical web-search tool fees applied separately per search call
Output $10.00 per 1 million output tokens
View model →

Accurate speech-to-text conversion, meeting transcription, call transcription, voice-agent input, and prompted domain-specific transcription

Type Other
Context 16K
Reasoning 2/10
Speed 8/10
Multimodal Audio input Streaming
Status

Deprecated; currently accessible; scheduled for API shutdown on 2027-02-26

Input $2.50 per 1M audio input tokens
Output $10.00 per 1M audio output tokens
View model →

Lower-cost multilingual speech transcription, meeting notes, call-center transcripts, voice-note conversion, and audio-to-text pipelines

Type Other
Context 16K
Reasoning 1/10
Speed 8/10
Multimodal Audio input Streaming
Status

Current and available

Input $1.25 per 1 million audio tokens
Output $5.00 per 1 million audio tokens
View model →
OpenAI logo
GPT-4o Mini

GPT-4o Mini TTS

Fast, controllable text-to-speech for narration, voice interfaces, customer service, accessibility, and realtime audio applications.

Type Other
Context 2K
Reasoning 1/10
Speed 9/10
Media output Streaming
Status

Current; the canonical alias currently points to the gpt-4o-mini-tts-2025-12-15 snapshot.

Input $0.60 per 1M text input tokens
Output $12.00 per 1M audio output tokens
View model →
OpenAI logo
GPT-4o Mini Search Preview

GPT-4o Mini Search Preview

Legacy Chat Completions applications requiring low-cost, search-grounded text responses

Type Lightweight
Context 128K
Reasoning 5/10
Speed 9/10
Tool use Web search Streaming
Status

Retired; access shut down on July 23, 2026

Input $0.15 per 1 million input tokens, plus a separate fee per web-search tool call
Output $0.60 per 1 million output tokens
View model →
OpenAI logo
GPT-4o Transcribe

GPT-4o Transcribe Diarize

Multi-speaker meeting, interview, call, podcast, and research transcription with speaker labels.

Type Other
Context 16K
Reasoning 2/10
Speed 8/10
Audio input Streaming
Status

Deprecated; currently accessible and scheduled for removal from the API on February 26, 2027.

Input $2.50 per 1M audio tokens
Output $10.00 per 1M audio tokens
View model →
OpenAI logo
GPT-5

GPT-5

Complex coding, reasoning, research, long-context analysis, visual understanding, tool-using agents, and structured professional workflows

Type Reasoning
Context 400K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Current canonical alias, but previous-generation model; dated snapshot gpt-5-2025-08-07 is deprecated and scheduled for API shutdown on 2026-12-11

Input $1.25 per 1 million input tokens; cached input $0.125 per 1 million tokens
Output $10.00 per 1 million output tokens
View model →

ChatGPT-aligned conversational applications, text generation, image-aware question answering, structured outputs, and tool-enabled workflows requiring GPT-5 compatibility

Type General Purpose
Context 128K
Reasoning 8/10
Speed 8/10
Multimodal Image input Tool use
Status

Deprecated

Input $1.25 per 1M input tokens; $0.125 per 1M cached input tokens
Output $10.00 per 1M output tokens
View model →

Cost-sensitive reasoning, coding assistance, structured extraction, document processing, high-volume automation, and tool-enabled workflows

Type Lightweight
Context 400K
Reasoning 8/10
Speed 9/10
Multimodal Image input Tool use
Status

Current API alias; dated snapshot gpt-5-mini-2025-08-07 is deprecated

Input US$0.25 per 1 million input tokens; cached input US$0.025 per 1 million tokens
Output US$2.00 per 1 million output tokens
View model →

High-volume classification, summarization, extraction, ranking, routing, image-assisted analysis, and lightweight coding subagents

Type Lightweight
Context 400K
Reasoning 7/10
Speed 10/10
Multimodal Image input Tool use
Status

Deprecated dated snapshot; currently accessible until scheduled shutdown on 2026-12-11

Input $0.05 per 1 million tokens; cached input $0.005 per 1 million tokens
Output $0.40 per 1 million tokens
View model →

Difficult research, mathematics, science, complex coding, high-stakes analysis, and tool-using workflows where maximum answer quality matters more than latency or cost.

Type Reasoning
Context 400K
Reasoning 10/10
Speed 3/10
Multimodal Image input Tool use
Status

Current canonical alias; dated snapshot gpt-5-pro-2025-10-06 is deprecated and scheduled for shutdown on 2026-12-11.

Input $15 per 1M input tokens
Output $120 per 1M output tokens
View model →

Agentic software engineering, repository-level coding, code review, debugging, refactoring, test generation, and frontend work using screenshots

Type Coding
Context 400K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Retired; API access shut down on 2026-07-23

Input $1.25 per 1M input tokens; $0.125 per 1M cached input tokens
Output $10.00 per 1M output tokens
View model →
OpenAI logo
GPT-5

GPT-5.1

Coding, long-context analysis, tool-using agents, structured outputs, and multi-step workflows

Type General Purpose
Context 400K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Current

Input $1.25 per 1 million input tokens; $0.125 per 1 million cached input tokens
Output $10.00 per 1 million output tokens
View model →

Conversational assistants, instruction following, image-grounded chat, structured extraction, streaming responses, and tool-using API workflows.

Type General Purpose
Context 128K
Reasoning 8/10
Speed 8/10
Multimodal Image input Tool use
Status

Retired; the API alias gpt-5.1-chat-latest was shut down on 2026-07-23. GPT-5.1 models were retired from ChatGPT on 2026-03-11.

Input $1.25 per 1 million input tokens; cached input $0.125 per 1 million tokens
Output $10.00 per 1 million output tokens
View model →

Agentic software engineering, code generation, debugging, refactoring, testing, code review, and long-running Codex workflows

Type Coding
Context 400K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Retired; API access shut down on July 23, 2026

Input $1.25 per 1M input tokens; $0.125 per 1M cached input tokens
Output $10.00 per 1M output tokens
View model →
OpenAI logo
GPT-5.1-Codex

GPT-5.1-Codex-Max

Long-running agentic coding, repository-scale refactoring, multi-file implementation, debugging, code review, pull-request creation, and extended Codex workflows.

Type Coding
Context 400K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Retired; API access ended 2026-07-23

Input $1.25 per 1M tokens; cached input $0.125 per 1M tokens
Output $10.00 per 1M tokens
View model →
OpenAI logo
GPT-5.1-Codex

GPT-5.1-Codex Mini

Cost-sensitive agentic coding, code editing, repository maintenance, and Codex-style workflows

Type Coding
Context 400K
Reasoning 7/10
Speed 9/10
Multimodal Image input Tool use
Status

Retired; API access shut down on 2026-07-23

Input $0.25 per 1 million tokens; cached input $0.025 per 1 million tokens
Output $2.00 per 1 million tokens
View model →
OpenAI logo
GPT-5.2

GPT-5.2

Complex professional work, long-context analysis, coding, document and spreadsheet workflows, visual understanding, and multi-step agents

Type Reasoning
Context 400K
Reasoning 9/10
Speed 7/10
Multimodal Image input Tool use
Status

Currently available; previous flagship model

Input $1.75 per 1M input tokens; $0.175 per 1M cached input tokens
Output $14.00 per 1M output tokens
View model →

ChatGPT-aligned conversational applications, general writing, summarization, translation, vision-enabled assistants, and tool-calling workflows

Type General Purpose
Context 128K
Reasoning 8/10
Speed 8/10
Multimodal Image input Tool use
Status

Retired; API access ended on 2026-08-10

Input $1.75 per 1 million input tokens; $0.175 per 1 million cached input tokens
Output $14.00 per 1 million output tokens
View model →

Complex professional reasoning, advanced analysis, scientific and mathematical work, high-quality coding, long-context document analysis, and tool-using workflows.

Type Reasoning
Context 400K
Reasoning 10/10
Speed 4/10
Multimodal Image input Tool use
Status

Previous Pro model; currently available through the Responses API

Input $21 per 1M tokens
Output $168 per 1M tokens
View model →

Long-horizon agentic coding, large refactors, code migrations, repository-scale changes, terminal workflows, Windows development and defensive cybersecurity

Type Coding
Context 400K
Reasoning 9/10
Speed 6/10
Multimodal Image input Tool use
Status

Retired; API access shut down on 2026-07-23

Input $1.75 per 1 million input tokens
Output $14.00 per 1 million output tokens
View model →

Fast general-purpose conversation, writing, summarization, text-and-image understanding, streaming responses, and function-calling applications

Type General Purpose
Context 128K
Reasoning 7/10
Speed 9/10
Multimodal Image input Tool use
Status

Retired; API access ended 2026-08-10

Input $1.75 per 1M tokens; cached input $0.175 per 1M tokens
Output $14.00 per 1M tokens
View model →

Long-running agentic software engineering, codebase maintenance, debugging, testing, web development, tool-driven development, and technical computer workflows

Type Coding
Context 400K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Current and available through OpenAI API and Codex surfaces

Input $1.75 per 1 million input tokens
Output $14.00 per 1 million output tokens
View model →
OpenAI logo
GPT-5.4

GPT-5.4

Complex professional work, advanced reasoning, software engineering, long-horizon agents, visual document analysis, computer use, research, and tool-heavy workflows

Type Reasoning
Context 1.05M
Reasoning 10/10
Speed 8/10
Multimodal Image input Tool use
Status

Current

Input $2.50 per 1 million input tokens; $0.25 per 1 million cached input tokens
Output $15.00 per 1 million output tokens
View model →

High-volume coding assistants, computer-use agents, subagents, tool calling, image reasoning, document workflows, and latency-sensitive applications

Type Lightweight
Context 400K
Reasoning 8/10
Speed 9/10
Multimodal Image input Tool use
Status

Current

Input $0.75 per 1 million input tokens; $0.075 per 1 million cached input tokens
Output $4.50 per 1 million output tokens
View model →

High-volume classification, data extraction, ranking, image understanding, routing, and lightweight coding subagents

Type Lightweight
Context 400K
Reasoning 7/10
Speed 9/10
Multimodal Image input Tool use
Status

Current; API-only model

Input $0.20 per 1M input tokens; $0.02 per 1M cached input tokens
Output $1.25 per 1M output tokens
View model →

High-stakes reasoning, professional knowledge work, long-context analysis, complex coding, web research and agentic workflows requiring maximum answer quality

Type Reasoning
Context 1.05M
Reasoning 10/10
Speed 4/10
Multimodal Image input Tool use
Status

Current; available in ChatGPT for Pro and Enterprise users and in the Responses API for developers

Input $30 per 1 million input tokens
Output $180 per 1 million output tokens
View model →

Authorized vulnerability research, defensive cybersecurity operations, malware analysis, security testing, and binary reverse engineering.

Type Coding
Reasoning 9/10
Status

Deprecated; currently accessible through restricted Trusted Access for Cyber channels; scheduled for API shutdown on October 1, 2026.

View model →
OpenAI logo
GPT-5.5

GPT-5.5

Complex coding, long-context research, professional analysis, tool-heavy agents, computer use, and multi-step workflow execution

Type General Purpose
Context 1.05M
Reasoning 10/10
Speed 8/10
Multimodal Image input Tool use
Status

Current; available through the OpenAI API, ChatGPT, and Codex

Input $5.00 per 1 million input tokens; cached input $0.50 per 1 million tokens
Output $30.00 per 1 million output tokens
View model →

High-accuracy reasoning, complex coding, long-context research, data analysis, and multi-step professional workflows

Type Reasoning
Context 1.05M
Reasoning 10/10
Speed 5/10
Multimodal Image input Tool use
Status

Current

Input USD 30 per 1 million input tokens; no cached-input discount
Output USD 180 per 1 million output tokens
View model →

Authorized vulnerability research, exploit validation, exploit-chain development, advanced security testing, vulnerability triage, and defensive cybersecurity agents

Type Coding
Context 400K
Reasoning 10/10
Speed 6/10
Multimodal Image input Tool use
Status

Current; restricted access through OpenAI Daybreak Red with separate approval and provisioning

Input $12.50 per 1 million input tokens; cached input $1.25 per 1 million tokens
Output $75.00 per 1 million output tokens
View model →

High-volume classification, summarization, routing, extraction, document understanding, agent automation, routine coding assistance, and cost-sensitive tool-using applications.

Type Lightweight
Context 1.05M
Reasoning 7/10
Speed 9/10
Multimodal Image input Tool use
Status

current

Input $0.20 per 1 million input tokens; cached input $0.02 per 1 million tokens; cache writes billed at 1.25x the uncached input rate. Requests with more than 272,000 input tokens are priced at 2x input for the full request.
Output $1.20 per 1 million output tokens. Requests with more than 272,000 input tokens are priced at 1.5x output for the full request.
View model →

Complex reasoning, coding, research, cybersecurity, science, long-context analysis, document-heavy workflows, and tool-using agents

Type Reasoning
Context 1.05M
Reasoning 10/10
Speed 8/10
Multimodal Image input Tool use
Status

Generally available

Input $4 per 1 million input tokens; cached input $0.40 per 1 million tokens
Output $20 per 1 million output tokens
View model →

Cost-conscious reasoning, coding agents, long-context analysis, structured business automation, research workflows, and tool-enabled production applications

Type General Purpose
Context 1.05M
Reasoning 8/10
Speed 8/10
Multimodal Image input Tool use
Status

Generally available

Input $2.00 per 1 million input tokens; $0.20 per 1 million cached input tokens
Output $12.00 per 1 million output tokens
View model →

Complex reasoning, agentic coding, computer use, web research, scientific and professional workflows, and long-context document tasks

Type Reasoning
Context 1.05M
Reasoning 10/10
Speed 8/10
Multimodal Image input Tool use
Status

Current; rolling out through the OpenAI API and selected ChatGPT, Azure, and Amazon Bedrock offerings

Input $10.00 per 1 million input tokens for Standard short-context processing; $1.00 per 1 million cached input tokens; $12.50 per 1 million cache-write tokens. Long-context input is $20.00 per 1 million tokens.
Output $50.00 per 1 million output tokens for Standard short-context processing; $75.00 per 1 million output tokens for long-context processing. Batch and Flex processing are priced at 50% of Standard rates.
View model →

High-volume reasoning, document analysis, coding assistance, retrieval-augmented generation, and repeatable agent workflows

Type Reasoning
Context 1.05M
Reasoning 8/10
Speed 9/10
Multimodal Image input Tool use
Status

Current and available

Input $0.10 per 1 million input tokens; $0.01 cached input; $0.125 cache writes. Long-context rates are $0.20 input, $0.02 cached input, and $0.25 cache writes per 1 million tokens.
Output $0.50 per 1 million output tokens for short context; $0.75 per 1 million output tokens for long context.
View model →

Complex coding, long-context reasoning, software engineering, research, computer use, and agentic workflows with tools.

Type Reasoning
Context 1.05M
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Current; generally available through the OpenAI API

Input USD 2.00 per 1M input tokens for short context; USD 4.00 per 1M input tokens for long context. Cached input is USD 0.20 short-context or USD 0.40 long-context per 1M tokens. Cache writes are USD 2.50 short-context or USD 5.00 long-context per 1M tokens un
Output USD 10.00 per 1M output tokens for short context; USD 15.00 per 1M output tokens for long context under Standard processing.
View model →
OpenAI logo
GPT-Audio

GPT-Audio

Audio-enabled chat applications, voice interfaces, spoken assistants, and applications requiring direct audio understanding and generation through Chat Completions.

Type Multimodal
Context 128K
Reasoning 5/10
Speed 7/10
Multimodal Audio input Media output
Status

Deprecated; scheduled for shutdown on January 20, 2027

Input $2.50 per 1M text tokens; $32.00 per 1M audio tokens
Output $10.00 per 1M text tokens; $64.00 per 1M audio tokens
View model →
OpenAI logo
GPT-Audio

GPT-Audio-1.5

Audio-in, audio-out conversational applications using the Chat Completions API, including voice assistants and tool-enabled spoken interfaces.

Type Multimodal
Context 128K
Reasoning 5/10
Speed 7/10
Multimodal Audio input Media output
Status

Current; generally available

Input Text: $2.50 per 1M tokens; audio: $32.00 per 1M audio tokens
Output Text: $10.00 per 1M tokens; audio: $64.00 per 1M audio tokens
View model →

Cost-sensitive, turn-based audio conversations, voice assistants, and audio-enabled applications using function calling

Type Multimodal
Context 128K
Reasoning 5/10
Speed 8/10
Multimodal Audio input Media output
Status

Deprecated; currently accessible but scheduled for API removal on 2027-01-20

Input $0.60 per 1 million text input tokens
Output $2.40 per 1 million text output tokens
View model →

Precise image editing, detailed creative work, high-fidelity generation, infographics, layouts, and workflows where fewer retries matter more than minimum latency

Type Multimodal
Speed 6/10
Multimodal Image input Media output
Status

Current

Input Text input: $5.00 per 1M tokens; image input: $8.00 per 1M tokens; cached text input: $1.25 per 1M tokens; cached image input: $2.00 per 1M tokens
Output Image output: $30.00 per 1M tokens
View model →
OpenAI logo
GPT-Live

GPT-Live 1

Natural low-latency voice agents, customer support, conversational workflows, live assistance, and applications requiring interruption-aware speech interaction

Type Multimodal
Reasoning 7/10
Speed 10/10
Multimodal Audio input Media output
Status

Current; available in the OpenAI API

Input $0.05 per voice-session minute, billed per second; backend model and tool usage billed separately
Output $0.05 per voice-session minute, billed per second; OpenAI documents this as the voice-session price rather than separate audio input and output token rates
View model →
OpenAI logo
GPT-Live-Transcribe

GPT-Live-Transcribe

Low-latency live captions, realtime call transcription, microphone streams, telephony audio, and voice-interface speech recognition

Type Other
Reasoning 1/10
Speed 9/10
Multimodal Audio input Streaming
Status

Current and generally available for realtime transcription

Input $0.017 per minute of realtime audio
Output Included in the per-minute realtime audio price; OpenAI does not list a separate output-token price
View model →

Local and private reasoning applications, coding assistants, agentic workflows, on-device or edge inference, fine-tuning, and cost-sensitive deployments with suitable hardware.

Type Reasoning
Context 131K
Reasoning 8/10
Speed 8/10
Tool use Web search Fine-tuning
Status

Current open-weight model; downloadable and usable through self-hosted or third-party inference infrastructure. Not served through the OpenAI API or ChatGPT.

Input No official OpenAI API input price; self-hosting and third-party hosting costs vary.
Output No official OpenAI API output price; self-hosting and third-party hosting costs vary.
View model →

Self-hosted reasoning, coding, agentic workflows, private deployments, research, and fine-tuning

Type Reasoning
Context 131K
Reasoning 8/10
Speed 6/10
Tool use Web search Fine-tuning
Status

Current open-weight model; downloadable and deployable locally or through third-party providers; not available through the OpenAI API

View model →
OpenAI logo
GPT-OSS-Safeguard

gpt-oss-safeguard-20b

Policy-based safety classification, LLM input/output filtering, content labeling, trust and safety review, and self-hosted moderation workflows

Type Other
Context 131K
Reasoning 8/10
Speed 6/10
Status

Research preview; currently available as an open-weight model

View model →
OpenAI logo
gpt-oss-safeguard

gpt-oss-safeguard-120b

Custom-policy safety classification, LLM input and output filtering, trust and safety labeling, nuanced moderation review, and offline safety analysis

Type Safety
Context 131K
Reasoning 8/10
Speed 3/10
Status

Research preview; open-weight and downloadable

View model →
OpenAI logo
GPT-Realtime

GPT-Realtime

Low-latency speech-to-speech voice agents, realtime customer support, education, accessibility, and conversational applications with function calling

Type Realtime
Context 32K
Reasoning 5/10
Speed 9/10
Multimodal Image input Audio input
Status

Deprecated; scheduled for API shutdown on January 20, 2027

Input Text: $4.00 per 1M tokens; cached text: $0.40 per 1M tokens; audio: $32.00 per 1M tokens; cached audio: $0.40 per 1M tokens; image: $5.00 per 1M tokens; cached image: $0.50 per 1M tokens
Output Text: $16.00 per 1M tokens; audio: $64.00 per 1M tokens
View model →
OpenAI logo
GPT-Realtime

GPT-Realtime-1.5

Low-latency speech-to-speech voice agents, customer support, realtime assistants, and audio applications that need function calling.

Type Realtime Audio
Context 32K
Reasoning 6/10
Speed 9/10
Multimodal Image input Audio input
Status

Active and currently available

Input $4.00 per 1M text tokens; $32.00 per 1M audio tokens; $5.00 per 1M image tokens. Cached input: $0.40 per 1M text or audio tokens and $0.50 per 1M image tokens.
Output $16.00 per 1M text tokens; $64.00 per 1M audio tokens.
View model →
OpenAI logo
GPT-Realtime

GPT-Realtime-2

Reasoning voice agents, speech-to-speech applications, customer support, live assistants, tool-driven workflows, and long conversational sessions

Type Multimodal
Context 128K
Reasoning 9/10
Speed 7/10
Multimodal Image input Audio input
Status

Current

Input Text: $4.00 per 1M tokens; cached text: $0.40 per 1M; audio: $32.00 per 1M tokens; cached audio: $0.40 per 1M; image: $5.00 per 1M tokens; cached image: $0.50 per 1M
Output Text: $24.00 per 1M tokens; audio: $64.00 per 1M tokens
View model →
OpenAI logo
GPT-Realtime

GPT-Realtime-2.1

Low-latency speech-to-speech agents, customer-service voice workflows, realtime tool use, telephony, and multimodal assistants with image input

Type Realtime
Context 128K
Reasoning 8/10
Speed 8/10
Multimodal Image input Audio input
Status

current

Input Text: $4.00 per 1M tokens; cached text: $0.40 per 1M; audio: $32.00 per 1M audio tokens; cached audio: $0.40 per 1M; image: $5.00 per 1M tokens; cached image: $0.50 per 1M
Output Text: $24.00 per 1M tokens; audio: $64.00 per 1M audio tokens
View model →

Low-latency spoken translation, multilingual calls, live interpretation, broadcasts, meetings, lessons, video rooms, captions, and translated audio experiences.

Type Other
Context 16K
Reasoning 3/10
Speed 9/10
Audio input Media output Streaming
Status

Current

Input $0.034 per minute of realtime audio
Output $0.034 per minute of realtime audio
View model →

Low-latency live transcription, captions, meeting notes, call analysis, voice-agent input, and continuous speech-to-text workflows

Type Other
Context 16K
Reasoning 1/10
Speed 9/10
Multimodal Audio input Streaming
Status

Current and available through the OpenAI Realtime API for realtime transcription

Input $0.017 per minute of audio
Output Included in the audio-duration transcription price; no separate text-output price documented
View model →
OpenAI logo
GPT-Realtime

GPT-Realtime Mini

Cost-sensitive realtime voice agents, speech-to-speech applications, interactive assistants, and multimodal interfaces

Type Realtime
Context 32K
Reasoning 5/10
Speed 8/10
Multimodal Image input Audio input
Status

Deprecated; currently accessible but scheduled for API shutdown on 2027-01-20

Input $0.60 per 1M text input tokens; $0.06 per 1M cached text input tokens
Output $2.40 per 1M text output tokens
View model →
OpenAI logo
GPT-Realtime-2.1

GPT-Realtime-2.1 Mini

Lower-cost, low-latency realtime voice agents, speech-to-speech assistants, and tool-enabled conversational applications

Type Lightweight
Context 128K
Reasoning 7/10
Speed 9/10
Multimodal Image input Audio input
Status

Current

Input Text: $0.60 per 1M tokens; cached text: $0.06 per 1M; audio: $10.00 per 1M tokens; cached audio: $0.30 per 1M; image: $0.80 per 1M tokens; cached image: $0.08 per 1M
Output Text: $2.40 per 1M tokens; audio: $20.00 per 1M tokens
View model →
OpenAI logo
GPT-Rosalind

GPT-Rosalind

Governed biology, genomics, medicinal chemistry, protein analysis, drug discovery, literature synthesis, wet-lab troubleshooting, and scientific tool workflows

Type Reasoning
Reasoning 9/10
Speed 5/10
Multimodal Image input Tool use
Status

Generally available to eligible organizations through the trusted-access program; approved internal life sciences research only

Input $5 per 1M input tokens; $0.50 per 1M cached input tokens
Output $25 per 1M output tokens
View model →
OpenAI logo
GPT-Transcribe

GPT-Transcribe

High-accuracy transcription of recorded audio, streamed file transcripts, multilingual recordings, and domain-specific speech with keyword or language hints

Type Other
Reasoning 2/10
Speed 8/10
Multimodal Audio input Streaming
Status

Current

Input $0.0045 per audio minute
Output No separate output-token price; included in the per-minute transcription price
View model →

Existing ChatGPT image-generation and image-editing integrations

Type Multimodal
Reasoning 1/10
Speed 7/10
Multimodal Image input Media output
Status

Deprecated; currently accessible; scheduled for shutdown on 2026-12-01

Input Text: $5.00 per 1M tokens; cached text: $1.25 per 1M tokens; image: $8.00 per 1M tokens; cached image: $2.00 per 1M tokens. Per-image generation: low $0.009-$0.013, medium $0.034-$0.05, high $0.133-$0.20 depending on size.
Output Text: $10.00 per 1M tokens; image: $32.00 per 1M tokens. Per-image generation: low $0.009-$0.013, medium $0.034-$0.05, high $0.133-$0.20 depending on size.
View model →
OpenAI logo
GPT Image

GPT-Image-1

API-based image generation, image editing, reference-image workflows, inpainting, marketing assets, e-commerce imagery, and visual content production

Type Image Generation
Speed 6/10
Multimodal Image input Media output
Status

Deprecated; currently accessible and scheduled to shut down on 2026-10-23

Input Text input: $5.00 per 1M tokens; cached text input: $1.25 per 1M tokens; image input: $10.00 per 1M image tokens; cached image input: $2.50 per 1M image tokens
Output Image generation per image: low $0.011 at 1024x1024 or $0.016 at 1024x1536 and 1536x1024; medium $0.042 or $0.063; high $0.167 or $0.25. Image output tokens: $40.00 per 1M tokens.
View model →
OpenAI logo
GPT Image

GPT-Image-1.5

Production image generation, image editing, branded graphics, ecommerce product imagery, marketing assets, and workflows requiring preservation of important visual details

Type Other
Reasoning 1/10
Speed 7/10
Multimodal Image input Media output
Status

Deprecated; currently accessible with API shutdown scheduled for 2026-12-01

Input $5.00 per 1M text tokens; $8.00 per 1M image tokens; cached input $1.25 per 1M text tokens and $2.00 per 1M image tokens
Output $10.00 per 1M text tokens; $32.00 per 1M image tokens; image generation $0.009-$0.20 per image depending on quality and resolution
View model →
OpenAI logo
GPT Image

GPT-Image-2

High-quality text-to-image generation, reference-based image editing, text-heavy visual assets, product imagery, marketing creatives, and production design workflows.

Type Multimodal
Speed 8/10
Multimodal Image input Media output
Status

Active; GPT Image 2.5 models are available for newer workflows, but GPT-Image-2 remains accessible as a documented API model.

Input $8.00 per 1M image input tokens; $2.00 per 1M cached image input tokens; $5.00 per 1M text input tokens; $1.25 per 1M cached text input tokens
Output $30.00 per 1M image output tokens; $10.00 per 1M text output tokens where applicable
View model →
OpenAI logo
GPT Image 1

GPT-Image-1 Mini

Cost-sensitive image generation and editing, high-volume variations, rapid ideation, previews, lightweight personalization, and draft creative assets.

Type Multimodal
Speed 8/10
Multimodal Image input Media output
Status

Deprecated; currently accessible but scheduled for API shutdown on 2026-12-01.

Input Text input: $2.00 per 1M tokens; cached text input: $0.20 per 1M tokens. Image input: $2.50 per 1M image tokens; cached image input: $0.25 per 1M image tokens.
Output Image output: $8.00 per 1M image tokens. Per-image generation: $0.005-$0.036 at 1024x1024 and $0.006-$0.052 at 1024x1536 or 1536x1024, depending on quality.
View model →
OpenAI logo
GPT Image 2.5

GPT-Image-2.5 Flare

Fast, high-quality image generation and editing, creator content, product experiences, visual search, rapid prototyping, and high-volume workflows

Type Multimodal
Speed 9/10
Multimodal Image input Media output
Status

Current; available through the OpenAI API

Input Text input: $5.00 per 1M tokens; cached text input: $1.25 per 1M tokens; image input: $8.00 per 1M image tokens; cached image input: $2.00 per 1M image tokens
Output Image output: $30.00 per 1M image tokens; text output is not billed because the model outputs images
View model →
OpenAI logo
o-series

o3

Complex reasoning, advanced coding, mathematics, science, technical research, visual analysis and multi-step tool workflows

Type Reasoning
Context 200K
Reasoning 9/10
Speed 6/10
Multimodal Image input Tool use
Status

Current canonical alias; o3-2025-04-16 snapshot deprecated and scheduled for API shutdown on December 11, 2026

Input $2.00 per 1M input tokens; $0.50 per 1M cached input tokens. Batch: $1.00 input and $0.25 cached input per 1M tokens.
Output $8.00 per 1M output tokens. Batch: $4.00 per 1M output tokens.
View model →
OpenAI logo
o1

o1

Complex reasoning, mathematics, science, coding analysis, visual reasoning, and high-accuracy multi-step tasks

Type Reasoning
Context 200K
Reasoning 9/10
Speed 4/10
Multimodal Image input Tool use
Status

Deprecated; still documented in the OpenAI API model catalog

Input $15.00 per 1M input tokens; $7.50 per 1M cached input tokens
Output $60.00 per 1M output tokens
View model →

Historically, difficult mathematics, science, coding, and other multi-step reasoning tasks requiring extended deliberation

Type Reasoning
Context 128K
Reasoning 9/10
Speed 4/10
Tool use Streaming
Status

Retired; API access shut down on 2025-07-28

Input $15.00 per 1 million input tokens; cached input $7.50 per 1 million tokens
Output $60.00 per 1 million output tokens
View model →

Cost-sensitive mathematics, science, algorithmic programming, debugging, and text-only reasoning

Type Reasoning
Context 128K
Reasoning 7/10
Speed 8/10
Streaming
Status

Deprecated

Input $1.10 per 1M input tokens; $0.55 per 1M cached input tokens
Output $4.40 per 1M output tokens
View model →

Complex reasoning, difficult technical analysis, advanced programming, research workflows, and tasks where answer consistency matters more than latency or cost.

Type Reasoning
Context 200K
Reasoning 9/10
Speed 3/10
Multimodal Image input Tool use
Status

Deprecated in OpenAI's current model catalog; the dated snapshot o1-pro-2025-03-19 is also marked deprecated. No exact shutdown date for the canonical o1-pro alias was found in the reviewed official documentation.

Input $150 per 1 million input tokens
Output $600 per 1 million output tokens
View model →

Complex multi-step research, source synthesis, legal and scientific analysis, market research, and large-scale internal-data investigation

Type Reasoning
Context 200K
Reasoning 10/10
Speed 3/10
Multimodal Image input Tool use
Status

Deprecated

Input $10.00 per 1M input tokens; $2.50 per 1M cached input tokens
Output $40.00 per 1M output tokens
View model →

Coding, mathematics, science, technical analysis, structured extraction, text-to-SQL, and multi-step reasoning

Type Reasoning
Context 200K
Reasoning 8/10
Speed 8/10
Tool use Streaming
Status

Current canonical alias with deprecated snapshot; o3-mini-2025-01-31 is scheduled for API shutdown on 2026-10-23

Input $1.10 per 1M input tokens; $0.55 per 1M cached input tokens
Output $4.40 per 1M output tokens
View model →

High-reliability reasoning, advanced mathematics, scientific analysis, complex coding, research, and multi-step professional work.

Type Reasoning
Context 200K
Reasoning 10/10
Speed 3/10
Multimodal Image input Tool use
Status

Current canonical alias; the dated snapshot o3-pro-2025-06-10 is marked deprecated in the model documentation.

Input $20 per 1 million input tokens
Output $80 per 1 million output tokens
View model →

Fast, cost-sensitive reasoning; coding; mathematics; visual analysis; structured extraction; high-volume tool-using agents

Type Reasoning
Context 200K
Reasoning 8/10
Speed 9/10
Multimodal Image input Tool use
Status

Deprecated; currently available through the API; scheduled for shutdown on 2026-10-23

Input $1.10 per 1 million input tokens; $0.275 per 1 million cached input tokens
Output $4.40 per 1 million output tokens
View model →

Complex multi-step research, source synthesis, market analysis, legal or scientific research, and long-form evidence-based reports.

Type Reasoning
Context 200K
Reasoning 9/10
Speed 8/10
Multimodal Image input Tool use
Status

Current canonical alias; the dated snapshot o4-mini-deep-research-2025-06-26 is deprecated.

Input $2.00 per 1M input tokens; $0.50 per 1M cached input tokens
Output $8.00 per 1M output tokens
View model →
OpenAI logo
omni-moderation

omni-moderation

Text and image safety classification, content filtering, AI-output screening, policy enforcement, and human-review routing

Type Moderation
Reasoning 2/10
Speed 8/10
Multimodal Image input
Status

Current

Input Free
Output Free
View model →
OpenAI logo
omni-moderation

omni-moderation-latest

Text and image safety classification, content filtering, moderation queues, policy enforcement, and generated-content screening

Type Other
Speed 9/10
Multimodal Image input
Status

Current default moderation model

Input Free through the Moderation API
Output Free through the Moderation API
View model →
OpenAI logo
Sora 2

Sora 2

Rapid video concepting, social clips, image-to-video experiments, prototypes, rough cuts, and audiovisual creative iteration

Type Multimodal
Reasoning 1/10
Speed 8/10
Multimodal Image input Media output
Status

Deprecated; currently accessible through the API as of September 23, 2026, with shutdown scheduled for September 24, 2026

Input Not token-priced; image and text inputs are included in video-generation requests
Output $0.10 per generated video second for 720x1280 portrait or 1280x720 landscape output
View model →

Production-quality text-to-video and image-guided video generation, cinematic prototypes, marketing assets, and high-resolution short clips with synchronized audio.

Type Video Generation
Multimodal Image input Media output
Status

Legacy; deprecated; currently accessible through September 23, 2026; scheduled for API shutdown on September 24, 2026

Input $0.30 per second at 720p; $0.50 per second at 1024p; $0.70 per second at 1080p. Batch pricing: $0.15, $0.25, and $0.35 per second respectively.
Output Video with synchronized audio; pricing is charged per generated second rather than per text or audio token.
View model →
OpenAI logo
text-embedding-3

text-embedding-3-large

High-quality semantic search, multilingual retrieval, RAG, recommendations, clustering, classification and similarity matching

Type Embedding
Context 8K
Reasoning 1/10
Speed 8/10
Status

Current and available through the OpenAI API

Input $0.13 per 1 million input tokens
Output No separate output-token charge; the model returns embedding vectors
View model →
OpenAI logo
text-embedding-3

text-embedding-3-small

Cost-efficient semantic search, retrieval-augmented generation, clustering, recommendations, anomaly detection, and text or code similarity

Type Embedding
Context 8K
Speed 9/10
Status

Current

Input $0.02 per 1 million input tokens
Output Not applicable; embedding output is billed through input-token usage
View model →
OpenAI logo
text-embedding-ada-002

text-embedding-ada-002

Legacy semantic search, retrieval, clustering, recommendations, anomaly detection, and classification systems already built around ada-002 vectors

Type Embedding
Context 8K
Reasoning 1/10
Speed 8/10
Status

Older embedding model; currently listed and accessible through the embeddings API

Input $0.10 per 1 million input tokens
View model →
OpenAI logo
text-moderation

text-moderation-stable

Legacy text-only safety classification and historical moderation integrations

Type Other
Reasoning 1/10
Speed 8/10
Status

Retired; access ended October 27, 2025

Input Free through the Moderation API
Output Free through the Moderation API
View model →
OpenAI logo
TTS-1

TTS-1

Low-latency text-to-speech, realtime-oriented voice interfaces, narration, accessibility, and automated audio generation

Type Other
Reasoning 1/10
Speed 9/10
Media output Streaming
Status

Current and accessible; optimized for low-latency text-to-speech

Input $15.00 per 1 million characters
View model →

High-quality text-to-speech generation, narration, accessibility audio, voice interfaces, and downloadable speech content

Type Other
Reasoning 1/10
Speed 7/10
Media output Streaming
Status

Current; available through the OpenAI Audio API speech endpoint

Input $30 per 1 million characters
Output Included in speech-generation pricing; output is billed by input characters rather than output tokens
View model →
OpenAI logo
Whisper

Whisper

Multilingual audio transcription, English speech translation, language identification, subtitles, captions, and word- or segment-level timestamps.

Type Other
Speed 8/10
Audio input Streaming
Status

Deprecated; currently accessible through the OpenAI API and scheduled for shutdown on February 26, 2027.

Input $0.006 per minute of audio
View model →