ChatGPT-style instant responses, general-purpose writing and analysis, image-aware conversations, and tool-assisted workflows
Type
General Purpose
Context
400K
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Current rolling alias; underlying model snapshot is regularly updated
Input
$5.00 per 1 million input tokens; cached input $0.50 per 1 million tokens
Output
$30.00 per 1 million output tokens
View model
→
Zero-shot image classification, image-text similarity, semantic image retrieval, multimodal indexing, and computer-vision research
Type
Multimodal
Context
77
Reasoning
2/10
Speed
7/10
Multimodal
Image input
Status
Public research release with downloadable weights; not verified as a current OpenAI hosted API model
View model
→
Codex CLI coding workflows, code question answering, code editing, repository tasks, and low-latency software-engineering assistance
Type
Coding
Context
200K
Reasoning
7/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Retired; API access ended on 2026-02-12
Input
$1.50 per 1M input tokens; $0.375 per 1M cached input tokens
Output
$6.00 per 1M output tokens
View model
→
Controlled browser automation, computer-use research, UI testing, and repetitive interface workflows
Type
Other
Context
8K
Reasoning
6/10
Speed
7/10
Multimodal
Image input
Tool use
Input
$3.00 per 1 million input tokens; tool-specific computer-use calls may incur separate fees
Output
$12.00 per 1 million output tokens
View model
→
Historical research on text-to-image generation, legacy image workflows, and comparisons with newer OpenAI image models.
Multimodal
Image input
Media output
Status
Retired; deprecated and removed from the OpenAI API on May 12, 2026.
View model
→
Historical text-to-image generation, concept art, illustration, visual ideation, marketing imagery, and prompt-following research
Type
Other
Reasoning
1/10
Speed
6/10
Media output
Status
Retired; deprecated and removed from the OpenAI API on May 12, 2026
Input
Not applicable to current use; historical pricing was charged per generated image rather than per input token
Output
Historical API pricing started at $0.04 per 1024×1024 standard-quality image; higher prices applied to HD and larger formats
View model
→
Maintaining legacy text-completion applications, historical GPT-3 base-model behavior, and existing compatible fine-tuned workflows before shutdown
Type
Lightweight
Reasoning
2/10
Speed
7/10
Status
Deprecated; currently accessible but scheduled to shut down on 2026-09-28
Input
$0.40 per 1 million input tokens
Output
$0.40 per 1 million output tokens
View model
→
Legacy text completion, code continuation, and inference from existing davinci-002 fine-tuned models before shutdown
Type
General Purpose
Reasoning
3/10
Speed
6/10
Status
Deprecated; API access scheduled to shut down on September 28, 2026
Input
$2.00 per 1M tokens
Output
$2.00 per 1M tokens
View model
→
Low-cost, high-volume text generation, summarization, classification, extraction, simple chatbots, and legacy API integrations
Type
General Purpose
Context
16K
Reasoning
4/10
Speed
8/10
Fine-tuning
Status
Deprecated; still available through the OpenAI API
Input
$0.50 per 1 million input tokens
Output
$1.50 per 1 million output tokens
View model
→
Maintaining established GPT-4 integrations, general-purpose text generation, analysis, writing, and coding workloads
Type
General Purpose
Context
8K
Reasoning
8/10
Speed
5/10
Multimodal
Image input
Tool use
Status
Legacy; older high-intelligence GPT model
Input
$30 per 1 million prompt tokens
Output
$60 per 1 million completion tokens
View model
→
Legacy high-context text and image analysis, function calling, JSON-mode workflows, and existing GPT-4 Turbo integrations
Type
Multimodal
Context
128K
Reasoning
7/10
Speed
6/10
Multimodal
Image input
Tool use
Status
Deprecated but currently accessible; scheduled for shutdown on October 23, 2026
Input
$10 per 1 million input tokens
Output
$30 per 1 million output tokens
View model
→
Historical long-context text generation, document analysis, structured text generation, and general-purpose assistant applications.
Type
General Purpose
Context
128K
Reasoning
7/10
Speed
7/10
Fine-tuning
Status
Retired; the gpt-4-turbo-preview alias pointed to gpt-4-0125-preview, which was shut down on 2026-03-26.
Input
$10 per 1 million tokens
Output
$30 per 1 million tokens
View model
→
Software engineering, long-context document analysis, precise instruction following, structured extraction, tool-enabled agents, and image understanding
Type
General Purpose
Context
1.05M
Reasoning
7/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current; default GPT-4.1 alias with gpt-4.1-2025-04-14 snapshot
Input
$2.00 per 1M input tokens; $0.50 per 1M cached input tokens
Output
$8.00 per 1M output tokens
View model
→
Fast, cost-efficient instruction following, coding assistance, image understanding, structured extraction, tool calling, and long-context API applications
Type
Lightweight
Context
1.05M
Reasoning
6/10
Speed
9/10
Multimodal
Image input
Tool use
Input
$0.40 per 1 million tokens; cached input $0.10 per 1 million tokens
Output
$1.60 per 1 million tokens
View model
→
High-volume, latency-sensitive classification, extraction, routing, summarization, lightweight assistants, image-assisted analysis, and simple tool-calling workflows
Type
Lightweight
Context
1.05M
Reasoning
4/10
Speed
10/10
Multimodal
Image input
Tool use
Status
Deprecated; currently accessible as of September 23, 2026; scheduled for shutdown on October 23, 2026
Input
$0.10 per 1M input tokens; $0.025 per 1M cached input tokens
Output
$0.40 per 1M output tokens
View model
→
Historical general-purpose writing, creative work, nuanced communication, image understanding, programming assistance, and applications needing function calling or structured outputs.
Type
General Purpose
Context
128K
Reasoning
7/10
Speed
3/10
Multimodal
Image input
Tool use
Status
Retired from the API on 2025-07-14; retired from ChatGPT in June 2026
Input
$75.00 per 1 million input tokens; $37.50 per 1 million cached input tokens
Output
$150.00 per 1 million output tokens
View model
→
Fast general-purpose conversations, vision, voice interactions, coding, and everyday productivity
Type
Multimodal
Context
128K
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Audio input
View model
→
General-purpose assistants, image understanding, coding help, structured extraction, multilingual generation, and latency-sensitive API workflows
Type
Multimodal
Context
128K
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Current in the OpenAI API; retired from ChatGPT on 2026-02-13. The gpt-4o-2024-05-13 snapshot is scheduled for API shutdown on 2026-10-23.
Input
$2.50 per 1M input tokens; $1.25 per 1M cached input tokens
Output
$10.00 per 1M output tokens
View model
→
Voice assistants, spoken conversational agents, audio-enabled customer service, and applications requiring direct audio understanding and speech generation
Type
Multimodal
Context
128K
Reasoning
7/10
Speed
6/10
Multimodal
Audio input
Media output
Status
Retired; API access ended May 7, 2026
Input
$2.50 per 1M text input tokens; $40 per 1M audio input tokens
Output
$10.00 per 1M text output tokens; $80 per 1M audio output tokens
View model
→
Low-cost, high-volume text and image understanding, classification, extraction, translation, tagging, customer support, routing, and structured data generation
Type
Lightweight
Context
128K
Reasoning
5/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Current canonical model alias; dated snapshot gpt-4o-mini-2024-07-18 is available
Input
$0.15 per 1M input tokens
Output
$0.60 per 1M output tokens
View model
→
Lower-cost audio understanding, conversational voice interfaces, and applications requiring text and spoken-audio input/output
Type
Multimodal
Context
128K
Reasoning
3/10
Speed
8/10
Multimodal
Audio input
Media output
Status
Deprecated; scheduled for API shutdown on 2027-01-20
Input
Text: $0.15 per 1M tokens; audio: $10.00 per 1M tokens
Output
Text: $0.60 per 1M tokens; audio: $20.00 per 1M tokens
View model
→
Low-cost realtime voice assistants, speech-to-speech interfaces, interactive audio applications, and conversational prototypes
Type
Lightweight
Context
16K
Reasoning
4/10
Speed
9/10
Multimodal
Audio input
Media output
Status
Deprecated; scheduled for API removal on 2027-01-20
Input
Text: $0.60 per 1M input tokens; audio: $10.00 per 1M audio tokens; cached input: $0.30 per 1M tokens
Output
Text: $2.40 per 1M output tokens; audio: $20.00 per 1M audio tokens
View model
→
Low-latency voice assistants, speech-to-speech applications, live translation, language learning, and interactive customer support
Type
Multimodal
Context
32K
Reasoning
6/10
Speed
9/10
Multimodal
Audio input
Media output
Status
Retired; API access ended 2026-05-07
Input
$5 per 1M text tokens; $40 per 1M audio tokens; cached text and audio input $2.50 per 1M tokens
Output
$20 per 1M text tokens; $80 per 1M audio tokens
View model
→
Historical web-search applications built around OpenAI Chat Completions
Type
Other
Context
128K
Reasoning
5/10
Speed
6/10
Tool use
Streaming
Status
Retired; shut down on 2026-07-23
Input
$2.50 per 1 million input tokens; historical web-search tool fees applied separately per search call
Output
$10.00 per 1 million output tokens
View model
→
Accurate speech-to-text conversion, meeting transcription, call transcription, voice-agent input, and prompted domain-specific transcription
Type
Other
Context
16K
Reasoning
2/10
Speed
8/10
Multimodal
Audio input
Streaming
Status
Deprecated; currently accessible; scheduled for API shutdown on 2027-02-26
Input
$2.50 per 1M audio input tokens
Output
$10.00 per 1M audio output tokens
View model
→
Lower-cost multilingual speech transcription, meeting notes, call-center transcripts, voice-note conversion, and audio-to-text pipelines
Type
Other
Context
16K
Reasoning
1/10
Speed
8/10
Multimodal
Audio input
Streaming
Status
Current and available
Input
$1.25 per 1 million audio tokens
Output
$5.00 per 1 million audio tokens
View model
→
Fast, controllable text-to-speech for narration, voice interfaces, customer service, accessibility, and realtime audio applications.
Type
Other
Context
2K
Reasoning
1/10
Speed
9/10
Media output
Streaming
Status
Current; the canonical alias currently points to the gpt-4o-mini-tts-2025-12-15 snapshot.
Input
$0.60 per 1M text input tokens
Output
$12.00 per 1M audio output tokens
View model
→
GPT-4o Mini Search Preview
Legacy Chat Completions applications requiring low-cost, search-grounded text responses
Type
Lightweight
Context
128K
Reasoning
5/10
Speed
9/10
Tool use
Web search
Streaming
Status
Retired; access shut down on July 23, 2026
Input
$0.15 per 1 million input tokens, plus a separate fee per web-search tool call
Output
$0.60 per 1 million output tokens
View model
→
Multi-speaker meeting, interview, call, podcast, and research transcription with speaker labels.
Type
Other
Context
16K
Reasoning
2/10
Speed
8/10
Audio input
Streaming
Status
Deprecated; currently accessible and scheduled for removal from the API on February 26, 2027.
Input
$2.50 per 1M audio tokens
Output
$10.00 per 1M audio tokens
View model
→
Complex coding, reasoning, research, long-context analysis, visual understanding, tool-using agents, and structured professional workflows
Type
Reasoning
Context
400K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current canonical alias, but previous-generation model; dated snapshot gpt-5-2025-08-07 is deprecated and scheduled for API shutdown on 2026-12-11
Input
$1.25 per 1 million input tokens; cached input $0.125 per 1 million tokens
Output
$10.00 per 1 million output tokens
View model
→
ChatGPT-aligned conversational applications, text generation, image-aware question answering, structured outputs, and tool-enabled workflows requiring GPT-5 compatibility
Type
General Purpose
Context
128K
Reasoning
8/10
Speed
8/10
Multimodal
Image input
Tool use
Input
$1.25 per 1M input tokens; $0.125 per 1M cached input tokens
Output
$10.00 per 1M output tokens
View model
→
Cost-sensitive reasoning, coding assistance, structured extraction, document processing, high-volume automation, and tool-enabled workflows
Type
Lightweight
Context
400K
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Current API alias; dated snapshot gpt-5-mini-2025-08-07 is deprecated
Input
US$0.25 per 1 million input tokens; cached input US$0.025 per 1 million tokens
Output
US$2.00 per 1 million output tokens
View model
→
High-volume classification, summarization, extraction, ranking, routing, image-assisted analysis, and lightweight coding subagents
Type
Lightweight
Context
400K
Reasoning
7/10
Speed
10/10
Multimodal
Image input
Tool use
Status
Deprecated dated snapshot; currently accessible until scheduled shutdown on 2026-12-11
Input
$0.05 per 1 million tokens; cached input $0.005 per 1 million tokens
Output
$0.40 per 1 million tokens
View model
→
Difficult research, mathematics, science, complex coding, high-stakes analysis, and tool-using workflows where maximum answer quality matters more than latency or cost.
Type
Reasoning
Context
400K
Reasoning
10/10
Speed
3/10
Multimodal
Image input
Tool use
Status
Current canonical alias; dated snapshot gpt-5-pro-2025-10-06 is deprecated and scheduled for shutdown on 2026-12-11.
Input
$15 per 1M input tokens
Output
$120 per 1M output tokens
View model
→
Agentic software engineering, repository-level coding, code review, debugging, refactoring, test generation, and frontend work using screenshots
Type
Coding
Context
400K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Retired; API access shut down on 2026-07-23
Input
$1.25 per 1M input tokens; $0.125 per 1M cached input tokens
Output
$10.00 per 1M output tokens
View model
→
Coding, long-context analysis, tool-using agents, structured outputs, and multi-step workflows
Type
General Purpose
Context
400K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Input
$1.25 per 1 million input tokens; $0.125 per 1 million cached input tokens
Output
$10.00 per 1 million output tokens
View model
→
Conversational assistants, instruction following, image-grounded chat, structured extraction, streaming responses, and tool-using API workflows.
Type
General Purpose
Context
128K
Reasoning
8/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Retired; the API alias gpt-5.1-chat-latest was shut down on 2026-07-23. GPT-5.1 models were retired from ChatGPT on 2026-03-11.
Input
$1.25 per 1 million input tokens; cached input $0.125 per 1 million tokens
Output
$10.00 per 1 million output tokens
View model
→
Agentic software engineering, code generation, debugging, refactoring, testing, code review, and long-running Codex workflows
Type
Coding
Context
400K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Retired; API access shut down on July 23, 2026
Input
$1.25 per 1M input tokens; $0.125 per 1M cached input tokens
Output
$10.00 per 1M output tokens
View model
→
Long-running agentic coding, repository-scale refactoring, multi-file implementation, debugging, code review, pull-request creation, and extended Codex workflows.
Type
Coding
Context
400K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Retired; API access ended 2026-07-23
Input
$1.25 per 1M tokens; cached input $0.125 per 1M tokens
Output
$10.00 per 1M tokens
View model
→
Cost-sensitive agentic coding, code editing, repository maintenance, and Codex-style workflows
Type
Coding
Context
400K
Reasoning
7/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Retired; API access shut down on 2026-07-23
Input
$0.25 per 1 million tokens; cached input $0.025 per 1 million tokens
Output
$2.00 per 1 million tokens
View model
→
Complex professional work, long-context analysis, coding, document and spreadsheet workflows, visual understanding, and multi-step agents
Type
Reasoning
Context
400K
Reasoning
9/10
Speed
7/10
Multimodal
Image input
Tool use
Status
Currently available; previous flagship model
Input
$1.75 per 1M input tokens; $0.175 per 1M cached input tokens
Output
$14.00 per 1M output tokens
View model
→
ChatGPT-aligned conversational applications, general writing, summarization, translation, vision-enabled assistants, and tool-calling workflows
Type
General Purpose
Context
128K
Reasoning
8/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Retired; API access ended on 2026-08-10
Input
$1.75 per 1 million input tokens; $0.175 per 1 million cached input tokens
Output
$14.00 per 1 million output tokens
View model
→
Complex professional reasoning, advanced analysis, scientific and mathematical work, high-quality coding, long-context document analysis, and tool-using workflows.
Type
Reasoning
Context
400K
Reasoning
10/10
Speed
4/10
Multimodal
Image input
Tool use
Status
Previous Pro model; currently available through the Responses API
Input
$21 per 1M tokens
Output
$168 per 1M tokens
View model
→
Long-horizon agentic coding, large refactors, code migrations, repository-scale changes, terminal workflows, Windows development and defensive cybersecurity
Type
Coding
Context
400K
Reasoning
9/10
Speed
6/10
Multimodal
Image input
Tool use
Status
Retired; API access shut down on 2026-07-23
Input
$1.75 per 1 million input tokens
Output
$14.00 per 1 million output tokens
View model
→
Fast general-purpose conversation, writing, summarization, text-and-image understanding, streaming responses, and function-calling applications
Type
General Purpose
Context
128K
Reasoning
7/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Retired; API access ended 2026-08-10
Input
$1.75 per 1M tokens; cached input $0.175 per 1M tokens
Output
$14.00 per 1M tokens
View model
→
Long-running agentic software engineering, codebase maintenance, debugging, testing, web development, tool-driven development, and technical computer workflows
Type
Coding
Context
400K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current and available through OpenAI API and Codex surfaces
Input
$1.75 per 1 million input tokens
Output
$14.00 per 1 million output tokens
View model
→
Complex professional work, advanced reasoning, software engineering, long-horizon agents, visual document analysis, computer use, research, and tool-heavy workflows
Type
Reasoning
Context
1.05M
Reasoning
10/10
Speed
8/10
Multimodal
Image input
Tool use
Input
$2.50 per 1 million input tokens; $0.25 per 1 million cached input tokens
Output
$15.00 per 1 million output tokens
View model
→
High-volume coding assistants, computer-use agents, subagents, tool calling, image reasoning, document workflows, and latency-sensitive applications
Type
Lightweight
Context
400K
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Tool use
Input
$0.75 per 1 million input tokens; $0.075 per 1 million cached input tokens
Output
$4.50 per 1 million output tokens
View model
→
High-volume classification, data extraction, ranking, image understanding, routing, and lightweight coding subagents
Type
Lightweight
Context
400K
Reasoning
7/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Current; API-only model
Input
$0.20 per 1M input tokens; $0.02 per 1M cached input tokens
Output
$1.25 per 1M output tokens
View model
→
High-stakes reasoning, professional knowledge work, long-context analysis, complex coding, web research and agentic workflows requiring maximum answer quality
Type
Reasoning
Context
1.05M
Reasoning
10/10
Speed
4/10
Multimodal
Image input
Tool use
Status
Current; available in ChatGPT for Pro and Enterprise users and in the Responses API for developers
Input
$30 per 1 million input tokens
Output
$180 per 1 million output tokens
View model
→
Authorized vulnerability research, defensive cybersecurity operations, malware analysis, security testing, and binary reverse engineering.
Type
Coding
Reasoning
9/10
Status
Deprecated; currently accessible through restricted Trusted Access for Cyber channels; scheduled for API shutdown on October 1, 2026.
View model
→
Complex coding, long-context research, professional analysis, tool-heavy agents, computer use, and multi-step workflow execution
Type
General Purpose
Context
1.05M
Reasoning
10/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current; available through the OpenAI API, ChatGPT, and Codex
Input
$5.00 per 1 million input tokens; cached input $0.50 per 1 million tokens
Output
$30.00 per 1 million output tokens
View model
→
High-accuracy reasoning, complex coding, long-context research, data analysis, and multi-step professional workflows
Type
Reasoning
Context
1.05M
Reasoning
10/10
Speed
5/10
Multimodal
Image input
Tool use
Input
USD 30 per 1 million input tokens; no cached-input discount
Output
USD 180 per 1 million output tokens
View model
→
Authorized vulnerability research, exploit validation, exploit-chain development, advanced security testing, vulnerability triage, and defensive cybersecurity agents
Type
Coding
Context
400K
Reasoning
10/10
Speed
6/10
Multimodal
Image input
Tool use
Status
Current; restricted access through OpenAI Daybreak Red with separate approval and provisioning
Input
$12.50 per 1 million input tokens; cached input $1.25 per 1 million tokens
Output
$75.00 per 1 million output tokens
View model
→
High-volume classification, summarization, routing, extraction, document understanding, agent automation, routine coding assistance, and cost-sensitive tool-using applications.
Type
Lightweight
Context
1.05M
Reasoning
7/10
Speed
9/10
Multimodal
Image input
Tool use
Input
$0.20 per 1 million input tokens; cached input $0.02 per 1 million tokens; cache writes billed at 1.25x the uncached input rate. Requests with more than 272,000 input tokens are priced at 2x input for the full request.
Output
$1.20 per 1 million output tokens. Requests with more than 272,000 input tokens are priced at 1.5x output for the full request.
View model
→
Complex reasoning, coding, research, cybersecurity, science, long-context analysis, document-heavy workflows, and tool-using agents
Type
Reasoning
Context
1.05M
Reasoning
10/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Generally available
Input
$4 per 1 million input tokens; cached input $0.40 per 1 million tokens
Output
$20 per 1 million output tokens
View model
→
Cost-conscious reasoning, coding agents, long-context analysis, structured business automation, research workflows, and tool-enabled production applications
Type
General Purpose
Context
1.05M
Reasoning
8/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Generally available
Input
$2.00 per 1 million input tokens; $0.20 per 1 million cached input tokens
Output
$12.00 per 1 million output tokens
View model
→
Complex reasoning, agentic coding, computer use, web research, scientific and professional workflows, and long-context document tasks
Type
Reasoning
Context
1.05M
Reasoning
10/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current; rolling out through the OpenAI API and selected ChatGPT, Azure, and Amazon Bedrock offerings
Input
$10.00 per 1 million input tokens for Standard short-context processing; $1.00 per 1 million cached input tokens; $12.50 per 1 million cache-write tokens. Long-context input is $20.00 per 1 million tokens.
Output
$50.00 per 1 million output tokens for Standard short-context processing; $75.00 per 1 million output tokens for long-context processing. Batch and Flex processing are priced at 50% of Standard rates.
View model
→
High-volume reasoning, document analysis, coding assistance, retrieval-augmented generation, and repeatable agent workflows
Type
Reasoning
Context
1.05M
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Current and available
Input
$0.10 per 1 million input tokens; $0.01 cached input; $0.125 cache writes. Long-context rates are $0.20 input, $0.02 cached input, and $0.25 cache writes per 1 million tokens.
Output
$0.50 per 1 million output tokens for short context; $0.75 per 1 million output tokens for long context.
View model
→
Complex coding, long-context reasoning, software engineering, research, computer use, and agentic workflows with tools.
Type
Reasoning
Context
1.05M
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current; generally available through the OpenAI API
Input
USD 2.00 per 1M input tokens for short context; USD 4.00 per 1M input tokens for long context. Cached input is USD 0.20 short-context or USD 0.40 long-context per 1M tokens. Cache writes are USD 2.50 short-context or USD 5.00 long-context per 1M tokens un
Output
USD 10.00 per 1M output tokens for short context; USD 15.00 per 1M output tokens for long context under Standard processing.
View model
→
Audio-enabled chat applications, voice interfaces, spoken assistants, and applications requiring direct audio understanding and generation through Chat Completions.
Type
Multimodal
Context
128K
Reasoning
5/10
Speed
7/10
Multimodal
Audio input
Media output
Status
Deprecated; scheduled for shutdown on January 20, 2027
Input
$2.50 per 1M text tokens; $32.00 per 1M audio tokens
Output
$10.00 per 1M text tokens; $64.00 per 1M audio tokens
View model
→
Audio-in, audio-out conversational applications using the Chat Completions API, including voice assistants and tool-enabled spoken interfaces.
Type
Multimodal
Context
128K
Reasoning
5/10
Speed
7/10
Multimodal
Audio input
Media output
Status
Current; generally available
Input
Text: $2.50 per 1M tokens; audio: $32.00 per 1M audio tokens
Output
Text: $10.00 per 1M tokens; audio: $64.00 per 1M audio tokens
View model
→
Cost-sensitive, turn-based audio conversations, voice assistants, and audio-enabled applications using function calling
Type
Multimodal
Context
128K
Reasoning
5/10
Speed
8/10
Multimodal
Audio input
Media output
Status
Deprecated; currently accessible but scheduled for API removal on 2027-01-20
Input
$0.60 per 1 million text input tokens
Output
$2.40 per 1 million text output tokens
View model
→
Precise image editing, detailed creative work, high-fidelity generation, infographics, layouts, and workflows where fewer retries matter more than minimum latency
Type
Multimodal
Speed
6/10
Multimodal
Image input
Media output
Input
Text input: $5.00 per 1M tokens; image input: $8.00 per 1M tokens; cached text input: $1.25 per 1M tokens; cached image input: $2.00 per 1M tokens
Output
Image output: $30.00 per 1M tokens
View model
→
Natural low-latency voice agents, customer support, conversational workflows, live assistance, and applications requiring interruption-aware speech interaction
Type
Multimodal
Reasoning
7/10
Speed
10/10
Multimodal
Audio input
Media output
Status
Current; available in the OpenAI API
Input
$0.05 per voice-session minute, billed per second; backend model and tool usage billed separately
Output
$0.05 per voice-session minute, billed per second; OpenAI documents this as the voice-session price rather than separate audio input and output token rates
View model
→
Low-latency live captions, realtime call transcription, microphone streams, telephony audio, and voice-interface speech recognition
Type
Other
Reasoning
1/10
Speed
9/10
Multimodal
Audio input
Streaming
Status
Current and generally available for realtime transcription
Input
$0.017 per minute of realtime audio
Output
Included in the per-minute realtime audio price; OpenAI does not list a separate output-token price
View model
→
Local and private reasoning applications, coding assistants, agentic workflows, on-device or edge inference, fine-tuning, and cost-sensitive deployments with suitable hardware.
Type
Reasoning
Context
131K
Reasoning
8/10
Speed
8/10
Tool use
Web search
Fine-tuning
Status
Current open-weight model; downloadable and usable through self-hosted or third-party inference infrastructure. Not served through the OpenAI API or ChatGPT.
Input
No official OpenAI API input price; self-hosting and third-party hosting costs vary.
Output
No official OpenAI API output price; self-hosting and third-party hosting costs vary.
View model
→
Self-hosted reasoning, coding, agentic workflows, private deployments, research, and fine-tuning
Type
Reasoning
Context
131K
Reasoning
8/10
Speed
6/10
Tool use
Web search
Fine-tuning
Status
Current open-weight model; downloadable and deployable locally or through third-party providers; not available through the OpenAI API
View model
→
Policy-based safety classification, LLM input/output filtering, content labeling, trust and safety review, and self-hosted moderation workflows
Type
Other
Context
131K
Reasoning
8/10
Speed
6/10
Status
Research preview; currently available as an open-weight model
View model
→
Custom-policy safety classification, LLM input and output filtering, trust and safety labeling, nuanced moderation review, and offline safety analysis
Type
Safety
Context
131K
Reasoning
8/10
Speed
3/10
Status
Research preview; open-weight and downloadable
View model
→
Low-latency speech-to-speech voice agents, realtime customer support, education, accessibility, and conversational applications with function calling
Type
Realtime
Context
32K
Reasoning
5/10
Speed
9/10
Multimodal
Image input
Audio input
Status
Deprecated; scheduled for API shutdown on January 20, 2027
Input
Text: $4.00 per 1M tokens; cached text: $0.40 per 1M tokens; audio: $32.00 per 1M tokens; cached audio: $0.40 per 1M tokens; image: $5.00 per 1M tokens; cached image: $0.50 per 1M tokens
Output
Text: $16.00 per 1M tokens; audio: $64.00 per 1M tokens
View model
→
Low-latency speech-to-speech voice agents, customer support, realtime assistants, and audio applications that need function calling.
Type
Realtime Audio
Context
32K
Reasoning
6/10
Speed
9/10
Multimodal
Image input
Audio input
Status
Active and currently available
Input
$4.00 per 1M text tokens; $32.00 per 1M audio tokens; $5.00 per 1M image tokens. Cached input: $0.40 per 1M text or audio tokens and $0.50 per 1M image tokens.
Output
$16.00 per 1M text tokens; $64.00 per 1M audio tokens.
View model
→
Reasoning voice agents, speech-to-speech applications, customer support, live assistants, tool-driven workflows, and long conversational sessions
Type
Multimodal
Context
128K
Reasoning
9/10
Speed
7/10
Multimodal
Image input
Audio input
Input
Text: $4.00 per 1M tokens; cached text: $0.40 per 1M; audio: $32.00 per 1M tokens; cached audio: $0.40 per 1M; image: $5.00 per 1M tokens; cached image: $0.50 per 1M
Output
Text: $24.00 per 1M tokens; audio: $64.00 per 1M tokens
View model
→
Low-latency speech-to-speech agents, customer-service voice workflows, realtime tool use, telephony, and multimodal assistants with image input
Type
Realtime
Context
128K
Reasoning
8/10
Speed
8/10
Multimodal
Image input
Audio input
Input
Text: $4.00 per 1M tokens; cached text: $0.40 per 1M; audio: $32.00 per 1M audio tokens; cached audio: $0.40 per 1M; image: $5.00 per 1M tokens; cached image: $0.50 per 1M
Output
Text: $24.00 per 1M tokens; audio: $64.00 per 1M audio tokens
View model
→
Low-latency spoken translation, multilingual calls, live interpretation, broadcasts, meetings, lessons, video rooms, captions, and translated audio experiences.
Type
Other
Context
16K
Reasoning
3/10
Speed
9/10
Audio input
Media output
Streaming
Input
$0.034 per minute of realtime audio
Output
$0.034 per minute of realtime audio
View model
→
Low-latency live transcription, captions, meeting notes, call analysis, voice-agent input, and continuous speech-to-text workflows
Type
Other
Context
16K
Reasoning
1/10
Speed
9/10
Multimodal
Audio input
Streaming
Status
Current and available through the OpenAI Realtime API for realtime transcription
Input
$0.017 per minute of audio
Output
Included in the audio-duration transcription price; no separate text-output price documented
View model
→
Cost-sensitive realtime voice agents, speech-to-speech applications, interactive assistants, and multimodal interfaces
Type
Realtime
Context
32K
Reasoning
5/10
Speed
8/10
Multimodal
Image input
Audio input
Status
Deprecated; currently accessible but scheduled for API shutdown on 2027-01-20
Input
$0.60 per 1M text input tokens; $0.06 per 1M cached text input tokens
Output
$2.40 per 1M text output tokens
View model
→
Lower-cost, low-latency realtime voice agents, speech-to-speech assistants, and tool-enabled conversational applications
Type
Lightweight
Context
128K
Reasoning
7/10
Speed
9/10
Multimodal
Image input
Audio input
Input
Text: $0.60 per 1M tokens; cached text: $0.06 per 1M; audio: $10.00 per 1M tokens; cached audio: $0.30 per 1M; image: $0.80 per 1M tokens; cached image: $0.08 per 1M
Output
Text: $2.40 per 1M tokens; audio: $20.00 per 1M tokens
View model
→
Governed biology, genomics, medicinal chemistry, protein analysis, drug discovery, literature synthesis, wet-lab troubleshooting, and scientific tool workflows
Type
Reasoning
Reasoning
9/10
Speed
5/10
Multimodal
Image input
Tool use
Status
Generally available to eligible organizations through the trusted-access program; approved internal life sciences research only
Input
$5 per 1M input tokens; $0.50 per 1M cached input tokens
Output
$25 per 1M output tokens
View model
→
High-accuracy transcription of recorded audio, streamed file transcripts, multilingual recordings, and domain-specific speech with keyword or language hints
Type
Other
Reasoning
2/10
Speed
8/10
Multimodal
Audio input
Streaming
Input
$0.0045 per audio minute
Output
No separate output-token price; included in the per-minute transcription price
View model
→
Existing ChatGPT image-generation and image-editing integrations
Type
Multimodal
Reasoning
1/10
Speed
7/10
Multimodal
Image input
Media output
Status
Deprecated; currently accessible; scheduled for shutdown on 2026-12-01
Input
Text: $5.00 per 1M tokens; cached text: $1.25 per 1M tokens; image: $8.00 per 1M tokens; cached image: $2.00 per 1M tokens. Per-image generation: low $0.009-$0.013, medium $0.034-$0.05, high $0.133-$0.20 depending on size.
Output
Text: $10.00 per 1M tokens; image: $32.00 per 1M tokens. Per-image generation: low $0.009-$0.013, medium $0.034-$0.05, high $0.133-$0.20 depending on size.
View model
→
API-based image generation, image editing, reference-image workflows, inpainting, marketing assets, e-commerce imagery, and visual content production
Type
Image Generation
Speed
6/10
Multimodal
Image input
Media output
Status
Deprecated; currently accessible and scheduled to shut down on 2026-10-23
Input
Text input: $5.00 per 1M tokens; cached text input: $1.25 per 1M tokens; image input: $10.00 per 1M image tokens; cached image input: $2.50 per 1M image tokens
Output
Image generation per image: low $0.011 at 1024x1024 or $0.016 at 1024x1536 and 1536x1024; medium $0.042 or $0.063; high $0.167 or $0.25. Image output tokens: $40.00 per 1M tokens.
View model
→
Production image generation, image editing, branded graphics, ecommerce product imagery, marketing assets, and workflows requiring preservation of important visual details
Type
Other
Reasoning
1/10
Speed
7/10
Multimodal
Image input
Media output
Status
Deprecated; currently accessible with API shutdown scheduled for 2026-12-01
Input
$5.00 per 1M text tokens; $8.00 per 1M image tokens; cached input $1.25 per 1M text tokens and $2.00 per 1M image tokens
Output
$10.00 per 1M text tokens; $32.00 per 1M image tokens; image generation $0.009-$0.20 per image depending on quality and resolution
View model
→
High-quality text-to-image generation, reference-based image editing, text-heavy visual assets, product imagery, marketing creatives, and production design workflows.
Type
Multimodal
Speed
8/10
Multimodal
Image input
Media output
Status
Active; GPT Image 2.5 models are available for newer workflows, but GPT-Image-2 remains accessible as a documented API model.
Input
$8.00 per 1M image input tokens; $2.00 per 1M cached image input tokens; $5.00 per 1M text input tokens; $1.25 per 1M cached text input tokens
Output
$30.00 per 1M image output tokens; $10.00 per 1M text output tokens where applicable
View model
→
Cost-sensitive image generation and editing, high-volume variations, rapid ideation, previews, lightweight personalization, and draft creative assets.
Type
Multimodal
Speed
8/10
Multimodal
Image input
Media output
Status
Deprecated; currently accessible but scheduled for API shutdown on 2026-12-01.
Input
Text input: $2.00 per 1M tokens; cached text input: $0.20 per 1M tokens. Image input: $2.50 per 1M image tokens; cached image input: $0.25 per 1M image tokens.
Output
Image output: $8.00 per 1M image tokens. Per-image generation: $0.005-$0.036 at 1024x1024 and $0.006-$0.052 at 1024x1536 or 1536x1024, depending on quality.
View model
→
Fast, high-quality image generation and editing, creator content, product experiences, visual search, rapid prototyping, and high-volume workflows
Type
Multimodal
Speed
9/10
Multimodal
Image input
Media output
Status
Current; available through the OpenAI API
Input
Text input: $5.00 per 1M tokens; cached text input: $1.25 per 1M tokens; image input: $8.00 per 1M image tokens; cached image input: $2.00 per 1M image tokens
Output
Image output: $30.00 per 1M image tokens; text output is not billed because the model outputs images
View model
→
Complex reasoning, advanced coding, mathematics, science, technical research, visual analysis and multi-step tool workflows
Type
Reasoning
Context
200K
Reasoning
9/10
Speed
6/10
Multimodal
Image input
Tool use
Status
Current canonical alias; o3-2025-04-16 snapshot deprecated and scheduled for API shutdown on December 11, 2026
Input
$2.00 per 1M input tokens; $0.50 per 1M cached input tokens. Batch: $1.00 input and $0.25 cached input per 1M tokens.
Output
$8.00 per 1M output tokens. Batch: $4.00 per 1M output tokens.
View model
→
Complex reasoning, mathematics, science, coding analysis, visual reasoning, and high-accuracy multi-step tasks
Type
Reasoning
Context
200K
Reasoning
9/10
Speed
4/10
Multimodal
Image input
Tool use
Status
Deprecated; still documented in the OpenAI API model catalog
Input
$15.00 per 1M input tokens; $7.50 per 1M cached input tokens
Output
$60.00 per 1M output tokens
View model
→
Historically, difficult mathematics, science, coding, and other multi-step reasoning tasks requiring extended deliberation
Type
Reasoning
Context
128K
Reasoning
9/10
Speed
4/10
Tool use
Streaming
Status
Retired; API access shut down on 2025-07-28
Input
$15.00 per 1 million input tokens; cached input $7.50 per 1 million tokens
Output
$60.00 per 1 million output tokens
View model
→
Cost-sensitive mathematics, science, algorithmic programming, debugging, and text-only reasoning
Type
Reasoning
Context
128K
Reasoning
7/10
Speed
8/10
Streaming
Input
$1.10 per 1M input tokens; $0.55 per 1M cached input tokens
Output
$4.40 per 1M output tokens
View model
→
Complex reasoning, difficult technical analysis, advanced programming, research workflows, and tasks where answer consistency matters more than latency or cost.
Type
Reasoning
Context
200K
Reasoning
9/10
Speed
3/10
Multimodal
Image input
Tool use
Status
Deprecated in OpenAI's current model catalog; the dated snapshot o1-pro-2025-03-19 is also marked deprecated. No exact shutdown date for the canonical o1-pro alias was found in the reviewed official documentation.
Input
$150 per 1 million input tokens
Output
$600 per 1 million output tokens
View model
→
Complex multi-step research, source synthesis, legal and scientific analysis, market research, and large-scale internal-data investigation
Type
Reasoning
Context
200K
Reasoning
10/10
Speed
3/10
Multimodal
Image input
Tool use
Input
$10.00 per 1M input tokens; $2.50 per 1M cached input tokens
Output
$40.00 per 1M output tokens
View model
→
Coding, mathematics, science, technical analysis, structured extraction, text-to-SQL, and multi-step reasoning
Type
Reasoning
Context
200K
Reasoning
8/10
Speed
8/10
Tool use
Streaming
Status
Current canonical alias with deprecated snapshot; o3-mini-2025-01-31 is scheduled for API shutdown on 2026-10-23
Input
$1.10 per 1M input tokens; $0.55 per 1M cached input tokens
Output
$4.40 per 1M output tokens
View model
→
High-reliability reasoning, advanced mathematics, scientific analysis, complex coding, research, and multi-step professional work.
Type
Reasoning
Context
200K
Reasoning
10/10
Speed
3/10
Multimodal
Image input
Tool use
Status
Current canonical alias; the dated snapshot o3-pro-2025-06-10 is marked deprecated in the model documentation.
Input
$20 per 1 million input tokens
Output
$80 per 1 million output tokens
View model
→
Fast, cost-sensitive reasoning; coding; mathematics; visual analysis; structured extraction; high-volume tool-using agents
Type
Reasoning
Context
200K
Reasoning
8/10
Speed
9/10
Multimodal
Image input
Tool use
Status
Deprecated; currently available through the API; scheduled for shutdown on 2026-10-23
Input
$1.10 per 1 million input tokens; $0.275 per 1 million cached input tokens
Output
$4.40 per 1 million output tokens
View model
→
Complex multi-step research, source synthesis, market analysis, legal or scientific research, and long-form evidence-based reports.
Type
Reasoning
Context
200K
Reasoning
9/10
Speed
8/10
Multimodal
Image input
Tool use
Status
Current canonical alias; the dated snapshot o4-mini-deep-research-2025-06-26 is deprecated.
Input
$2.00 per 1M input tokens; $0.50 per 1M cached input tokens
Output
$8.00 per 1M output tokens
View model
→
Text and image safety classification, content filtering, AI-output screening, policy enforcement, and human-review routing
Type
Moderation
Reasoning
2/10
Speed
8/10
Multimodal
Image input
View model
→
Text and image safety classification, content filtering, moderation queues, policy enforcement, and generated-content screening
Multimodal
Image input
Status
Current default moderation model
Input
Free through the Moderation API
Output
Free through the Moderation API
View model
→
Rapid video concepting, social clips, image-to-video experiments, prototypes, rough cuts, and audiovisual creative iteration
Type
Multimodal
Reasoning
1/10
Speed
8/10
Multimodal
Image input
Media output
Status
Deprecated; currently accessible through the API as of September 23, 2026, with shutdown scheduled for September 24, 2026
Input
Not token-priced; image and text inputs are included in video-generation requests
Output
$0.10 per generated video second for 720x1280 portrait or 1280x720 landscape output
View model
→
Production-quality text-to-video and image-guided video generation, cinematic prototypes, marketing assets, and high-resolution short clips with synchronized audio.
Multimodal
Image input
Media output
Status
Legacy; deprecated; currently accessible through September 23, 2026; scheduled for API shutdown on September 24, 2026
Input
$0.30 per second at 720p; $0.50 per second at 1024p; $0.70 per second at 1080p. Batch pricing: $0.15, $0.25, and $0.35 per second respectively.
Output
Video with synchronized audio; pricing is charged per generated second rather than per text or audio token.
View model
→
High-quality semantic search, multilingual retrieval, RAG, recommendations, clustering, classification and similarity matching
Type
Embedding
Context
8K
Reasoning
1/10
Speed
8/10
Status
Current and available through the OpenAI API
Input
$0.13 per 1 million input tokens
Output
No separate output-token charge; the model returns embedding vectors
View model
→
Cost-efficient semantic search, retrieval-augmented generation, clustering, recommendations, anomaly detection, and text or code similarity
Type
Embedding
Context
8K
Speed
9/10
Input
$0.02 per 1 million input tokens
Output
Not applicable; embedding output is billed through input-token usage
View model
→
Legacy semantic search, retrieval, clustering, recommendations, anomaly detection, and classification systems already built around ada-002 vectors
Type
Embedding
Context
8K
Reasoning
1/10
Speed
8/10
Status
Older embedding model; currently listed and accessible through the embeddings API
Input
$0.10 per 1 million input tokens
View model
→
Legacy text-only safety classification and historical moderation integrations
Type
Other
Reasoning
1/10
Speed
8/10
Status
Retired; access ended October 27, 2025
Input
Free through the Moderation API
Output
Free through the Moderation API
View model
→
Low-latency text-to-speech, realtime-oriented voice interfaces, narration, accessibility, and automated audio generation
Type
Other
Reasoning
1/10
Speed
9/10
Media output
Streaming
Status
Current and accessible; optimized for low-latency text-to-speech
Input
$15.00 per 1 million characters
View model
→
High-quality text-to-speech generation, narration, accessibility audio, voice interfaces, and downloadable speech content
Type
Other
Reasoning
1/10
Speed
7/10
Media output
Streaming
Status
Current; available through the OpenAI Audio API speech endpoint
Input
$30 per 1 million characters
Output
Included in speech-generation pricing; output is billed by input characters rather than output tokens
View model
→
Multilingual audio transcription, English speech translation, language identification, subtitles, captions, and word- or segment-level timestamps.
Audio input
Streaming
Status
Deprecated; currently accessible through the OpenAI API and scheduled for shutdown on February 26, 2027.
Input
$0.006 per minute of audio
View model
→