GPT-5.6

GPT-5.6 Terra

by OpenAI · Generally available

GPT-5.6 Terra is OpenAI's mid-tier GPT-5.6 API model for balancing reasoning capability, long-context processing, coding, tool use and cost. It accepts text and image input, generates text, supports up to 1.05 million context tokens and 128,000 output tokens, and offers configurable reasoning, structured outputs, web search, code execution, computer use, batch processing and prompt caching.

Text Reasoning Coding
GPT-5.6 Terra is OpenAI's mid-tier GPT-5.6 model for applications that need more capability than a low-cost, speed-focused model but do not require the highest-priced flagship option. It combines reasoning controls, coding support, image understanding, long-context processing, structured outputs, and Responses API tools with a lower per-token price than GPT-5.6 Sol. The model is generally available through the OpenAI API under the model ID gpt-5.6-terra.
Outputs

What GPT-5.6 Terra can produce

Text
Inputs

What it can understand

Text Images Multimodal input
Capabilities

Supported features

Tool use Web search Streaming Structured output Prompt caching Batch API
Model profile

Performance characteristics

8/10 Reasoning
8/10 Coding
8/10 Speed
8/10 Cost efficiency
Specifications

Technical details

Model family GPT-5.6
Model type General Purpose
Context window 1.05M tokens
Maximum output 128K tokens
Knowledge cutoff 2026-02-16
Release date 2026-07-09
Status Generally available
Knowledge cutoff notes

The official model page identifies February 16, 2026 as the knowledge cutoff. This cutoff describes the underlying model and is not changed by web search, retrieval, tools, or user-provided context.

Model notes

The canonical API model ID is gpt-5.6-terra. OpenAI positions Terra as the GPT-5.6 model that balances intelligence and cost and describes it as roughly corresponding to the mini tier in earlier GPT-5 families. The model supports reasoning effort levels none, low, medium, high, xhigh, and max, with medium as the default. Documented Responses API tools include web search, file search, image generation, code interpreter, hosted shell, apply patch, skills, computer use, MCP, and tool search. The current model page lists $2.00 input and $12.00 output pricing; cached input is $0.20 per million tokens and cache writes are billed at 1.25 times the uncached input rate. Prompts exceeding 272,000 input tokens receive higher pricing for the full request. Image generation is available as a supported tool, but the model itself is documented as producing text output rather than native image output. Editorial scores are comparative estimates, not provider-published ratings.

Cost

Model pricing

Input $2.00 per 1 million input tokens; $0.20 per 1 million cached input tokens
Output $12.00 per 1 million output tokens
Model guide

GPT-5.6 Terra: Context, Pricing, Capabilities and API Support

GPT-5.6 Terra is OpenAI's balanced GPT-5.6 model for production applications that need strong reasoning, coding, long-context processing, and tool use without the flagship cost of GPT-5.6 Sol. It accepts text and image inputs, generates text, supports up to 1.05 million input tokens and 128,000 output tokens, and is priced at $2.00 per million input tokens and $12.00 per million output tokens.

GPT-5.6 Terra is OpenAI's balanced model in the GPT-5.6 family. Its purpose is to provide a practical middle ground between capability, response speed, and operating cost. It is aimed at production software, coding agents, research workflows, document analysis, and business automation where a model must handle substantial reasoning and tool use but does not need the most expensive model in the family.

Terra is not a general-purpose consumer product by itself. It is an API model that developers can call using its canonical model ID, gpt-5.6-terra. The specifications and prices below describe the supplied OpenAI model documentation and catalog information. Comparative comments about value, speed, or suitability are editorial evaluations based on those specifications rather than provider-published ratings.

What is GPT-5.6 Terra?

GPT-5.6 Terra is a general-purpose language model from OpenAI with support for reasoning, coding, structured responses, and tool-enabled workflows. OpenAI positions it between the flagship GPT-5.6 Sol and the faster, lower-cost GPT-5.6 Luna. The supplied research describes Terra as roughly comparable to the mini tier used in earlier GPT-5 families.

That positioning matters because Terra is designed for workloads where the cheapest available model may produce insufficiently reliable reasoning, while the flagship model would make every request unnecessarily expensive. Typical examples include an agent that reads long technical documents, a coding assistant that must use external tools, or an internal automation system that produces consistent structured records.

Technical specifications and limits

Terra has a documented context window of 1,050,000 tokens. A context window is the amount of input material the model can consider in one request, including instructions, conversation history, retrieved documents, tool results, and other supplied content. A context size of 1.05 million tokens is suitable for very large document collections and extended multi-step workflows, although sending very large requests can increase cost.

The maximum output is 128,000 tokens. This is an upper limit rather than a recommended size for every response. Most ordinary answers, code changes, and structured records will use much less, but the limit gives developers room for long analyses or substantial generated content.

SpecificationGPT-5.6 Terra
ProviderOpenAI
API model IDgpt-5.6-terra
AvailabilityGenerally available
Release dateJuly 9, 2026
Context window1,050,000 tokens
Maximum output128,000 tokens
Knowledge cutoffFebruary 16, 2026
Input typesText and images
Output typeText

The knowledge cutoff is separate from tool access. Terra's underlying training knowledge is documented through February 16, 2026. If an application uses web search or another retrieval system, the model can process newer information supplied through that tool, but this does not change the model's underlying cutoff.

Pricing and cost considerations

The supplied OpenAI model catalog lists GPT-5.6 Terra at $2.00 per million input tokens and $12.00 per million output tokens. Cached input tokens are listed at $0.20 per million tokens. Cache writes are billed at 1.25 times the uncached input rate. Requests containing more than 272,000 input tokens receive higher pricing for the full request.

Token categoryListed price
Input tokens$2.00 per 1 million
Cached input tokens$0.20 per 1 million
Output tokens$12.00 per 1 million
Cache writes1.25 times the uncached input rate

These are usage-based API prices, not a monthly subscription fee. The difference between input and output pricing is important for applications that generate long responses: output tokens cost substantially more than ordinary input tokens. Reusing stable instructions or reference material through prompt caching can reduce the cost of repeated requests, while very large prompts need special attention because the higher-pricing threshold applies above 272,000 input tokens.

Compared with GPT-5.6 Sol, Terra is the cost-conscious choice within the same family. Compared with a smaller or speed-oriented model such as GPT-5.6 Luna, Terra is better suited to tasks where deeper reasoning, longer context, or more capable tool use justifies additional spending. The supplied material does not provide benchmark figures, so these are positioning and cost trade-offs rather than measured performance claims.

Reasoning and coding capabilities

Terra supports configurable reasoning effort levels: none, low, medium, high, xhigh, and max. Medium is the documented default. Reasoning effort controls how much deliberate processing the model applies to a request. Lower settings can be appropriate for straightforward classification, extraction, or short answers; higher settings are more relevant to complex analysis, planning, debugging, and multi-step problem solving.

This flexibility helps developers balance latency and cost against answer quality. A production workflow might use a lower setting for routine records and increase the setting only when a task involves ambiguous requirements, difficult code, or several dependent decisions. The supplied research does not give benchmark scores for reasoning or coding, so Terra should not be described as having a verified numerical advantage over other models.

For coding applications, Terra can generate and explain code, analyze technical material, work through debugging tasks, and participate in tool-enabled coding agents. Its support for function calling, structured outputs, code interpretation, hosted shell access, and apply-patch workflows makes it suitable for applications that need more than plain text completion. Developers should still test generated code and apply normal security controls before allowing an agent to change files, execute commands, or access production systems.

Input, output, and tool support

GPT-5.6 Terra accepts text and image input and produces text output. It is therefore multimodal on the input side: an application can provide an image alongside instructions for analysis. It is not documented as a native image, audio, video, speech, music, or embedding-output model. Image generation can be accessed as a supported tool, but that is different from Terra directly producing native image output.

The documented capability set includes streaming, function calling, structured outputs, batch processing, and prompt caching. Structured outputs are useful when an application needs predictable fields, such as an invoice record, a classification result, or a machine-readable workflow decision. Function calling allows the model to request application-defined operations, while streaming lets an application display or process generated text as it arrives.

The supplied Responses API tool list includes web search, file search, code interpreter, hosted shell, computer use, MCP, tool search, image-generation tool access, apply patch, and skills. Tool availability and safe use depend on the surrounding implementation. A tool-enabled model does not independently grant access to a user's files, computer, databases, or external services; the application must provide and control those connections.

Main strengths and limitations

Where Terra is strongest

  • Balanced cost and capability: It is positioned below the flagship GPT-5.6 Sol in cost while retaining substantial reasoning and tool support.
  • Very long context: The 1.05-million-token window supports large document, repository, and research workflows.
  • Adjustable reasoning: Six reasoning effort levels allow developers to tune the trade-off between deliberation, latency, and spending.
  • Production-oriented interfaces: Streaming, function calling, structured outputs, batch processing, and caching support application development beyond simple chat.
  • Image understanding: Text-and-image input allows the model to analyze visual material while keeping text as its output format.

Where Terra is not the right fit

  • It is not a native media generator: Terra does not directly produce image, audio, video, speech, music, or embedding outputs. Image generation is available through a tool rather than as the model's native output mode.
  • No documented fine-tuning: The supplied model information lists fine-tuning as unsupported, so applications that require provider-hosted customization should consider another approach.
  • High-end tasks may favor the flagship: If maximum capability is more important than cost, GPT-5.6 Sol may be the more appropriate sibling option.
  • Simple, high-volume tasks may not need Terra: For short classification, basic extraction, or latency-sensitive requests, a smaller or faster model may offer better economics.
  • Tool use adds operational risk: Computer use, shell access, file operations, and external tools require permissions, monitoring, validation, and safeguards.

Best use cases for GPT-5.6 Terra

Terra is a good fit when the task combines meaningful reasoning with enough volume or duration that flagship pricing would be difficult to justify. Practical examples include:

  • Long-context analysis of contracts, technical manuals, research collections, or internal policies.
  • Coding agents that inspect repositories, plan changes, run tools, and return structured results.
  • Research assistants that combine document retrieval, web search, synthesis, and citations or evidence records.
  • Business automation that extracts information from text and images into a fixed schema.
  • Multi-step workflows that require function calls, code execution, hosted shell operations, or computer-use actions.
  • Batch processing where reasoning quality matters but every item does not require the highest-priced model.

When to choose this model

Choose GPT-5.6 Terra when you need a strong general-purpose API model with long context, adjustable reasoning, coding ability, image understanding, and broad tool support, but need to control per-request cost. It is especially suitable when the same system must handle both routine production requests and more demanding analysis without switching constantly between unrelated models.

Choose GPT-5.6 Sol instead when the application consistently prioritizes the highest capability available within the GPT-5.6 family and can accept higher pricing. Choose GPT-5.6 Luna or another speed- and cost-focused option when requests are short, predictable, and relatively simple, or when latency and request volume matter more than maximum reasoning depth. Choose a specialized image, audio, video, speech, or embedding model when native non-text output or vector generation is the central requirement.

Before deployment, test Terra with the application's real prompts, document sizes, tool calls, and required output schemas. Pay particular attention to output-token usage, requests above the 272,000-token pricing threshold, the effect of reasoning settings on latency and cost, and the safeguards needed for tools that can affect external systems.


Answers to Frequently Asked Questions

When should developers choose GPT-5.6 Terra?
Developers should choose GPT-5.6 Terra when they need long-context processing, adjustable reasoning, coding ability, image understanding, and tool use while controlling costs. It is well suited to document analysis, coding agents, research assistants, structured business automation, and multi-step workflows. Simpler or latency-sensitive tasks may be better served by GPT-5.6 Luna, while maximum-capability workloads may justify GPT-5.6 Sol.
What tools and capabilities does GPT-5.6 Terra support?
GPT-5.6 Terra supports reasoning, coding, streaming, function calling, structured outputs, batch processing, prompt caching, and image understanding. Its documented tool support includes web search, file search, code interpreter, hosted shell, computer use, MCP, tool search, image-generation access, apply patch, and skills.
How much does GPT-5.6 Terra cost?
GPT-5.6 Terra is listed at $2.00 per million input tokens, $12.00 per million output tokens, and $0.20 per million cached input tokens. Cache writes cost 1.25 times the uncached input rate. Requests with more than 272,000 input tokens receive higher pricing for the full request.
What is GPT-5.6 Terra?
GPT-5.6 Terra is OpenAI’s balanced API model for production software, coding agents, research workflows, document analysis, and business automation. It is positioned between the more capable GPT-5.6 Sol and the faster, lower-cost GPT-5.6 Luna.
What are GPT-5.6 Terra’s context window and output limits?
GPT-5.6 Terra has a context window of 1,050,000 tokens and a maximum output limit of 128,000 tokens. It accepts text and image inputs and produces text output.


Sources 4
Provider

About OpenAI