Claude Sonnet 5.5

Claude Sonnet 5.5

by Claude · Active (latest)

Claude Sonnet 5.5 is Anthropic’s active model for coding, long-context analysis, document creation, image understanding, and tool-enabled applications. It supports a 1 million-token context window, 128,000-token standard output, adaptive thinking, prompt caching, batch processing, and text-and-image input with text-only output.

Text Reasoning Coding
Claude Sonnet 5.5 is Anthropic’s September 2026 Sonnet model for users who need strong coding and reasoning without moving to the higher-cost Opus tier. It accepts text and images, produces text, supports adaptive thinking and tools, and is designed for software development, document workflows, long-context analysis, and agentic applications.
Outputs

What Claude Sonnet 5.5 can produce

Text
Inputs

What it can understand

Text Images Multimodal input
Capabilities

Supported features

Tool use Web search Streaming Structured output Prompt caching Batch API
Model profile

Performance characteristics

9/10 Reasoning
9/10 Coding
9/10 Speed
8/10 Cost efficiency
Specifications

Technical details

Model family Claude Sonnet 5.5
Model type General Purpose
Context window 1M tokens
Maximum output 128K tokens
Knowledge cutoff June 2026
Release date 2026-09-28
Status Active (latest)
Shutdown date 2027-09-28
Knowledge cutoff notes

Anthropic's model overview lists both the reliable knowledge cutoff and training data cutoff as June 2026.

Model notes

Canonical Claude API model ID is claude-sonnet-5-5. Amazon Bedrock uses anthropic.claude-sonnet-5-5. The model accepts text and images and returns text. It uses adaptive thinking and supports configurable effort, prompt caching, Files API, PDFs, vision, structured-output workflows, streaming, server-side tools, client-side tools, and batch processing. Five-minute cache writes cost $2.50 per million tokens, one-hour cache writes cost $4 per million tokens, and cache reads cost $0.20 per million tokens. Batch input and output receive a 50% discount. Batch output can reach 300,000 tokens with the applicable beta header. Non-default temperature, top_p, and top_k values return errors. Computer use on the Claude API and Google Cloud requires computer_toolset_20260801 rather than computer_20251124. Editorial scores are comparative estimates based on Anthropic's published capability, latency, coding, and knowledge-work information, not vendor ratings.

Cost

Model pricing

Input $2 per million tokens
Output $10 per million tokens
Model guide

Claude Sonnet 5.5: Anthropic’s Fast Model for Coding and Long-Context Work

Claude Sonnet 5.5 is Anthropic’s active general-purpose model for coding, long-context analysis, document creation, image understanding, and tool-enabled applications. It combines a 1 million-token context window, a 128,000-token standard output limit, adaptive thinking, vision, prompt caching, batch processing, and fast latency at a lower price than Claude Opus 5.5.

What is Claude Sonnet 5.5?

Claude Sonnet 5.5 is Anthropic’s active general-purpose language model for applications that need a combination of coding ability, reasoning, speed, and operational flexibility. Anthropic positions it between the higher-capability Claude Opus 5.5 and the lower-cost Haiku line. That makes Sonnet 5.5 a practical middle option: it is intended to handle demanding everyday work while remaining faster and less expensive than the provider’s top-tier model.

The model is aimed at software development, codebase maintenance, long-running coding tasks, professional documents, spreadsheets, presentations, image understanding, and tool-based workflows. It can analyze large amounts of material in one request and can work with external tools when an application supplies them.

According to Anthropic’s supplied model documentation, Claude Sonnet 5.5 was released on September 28, 2026. Its primary Claude API identifier is claude-sonnet-5-5. It is also available through Amazon Bedrock, Google Cloud, Microsoft Foundry, and Claude Platform on AWS. On Amazon Bedrock, the model identifier is anthropic.claude-sonnet-5-5.

Specifications at a glance

SpecificationClaude Sonnet 5.5
ProviderAnthropic
StatusActive and listed as latest
Context window1,000,000 tokens
Standard maximum output128,000 tokens
Batch maximum outputUp to 300,000 tokens with the applicable beta configuration
Knowledge cutoffJune 2026
InputText and images
OutputText
ReasoningAdaptive thinking with configurable effort
Latency positioningFast, according to Anthropic’s model information

The context window is the amount of information the model can consider in a request and its surrounding conversation. A 1 million-token limit is useful for large codebases, lengthy document collections, transcripts, and other tasks where splitting the source material into many smaller requests would be inconvenient. The context limit should not be confused with the maximum response size: a normal request can return up to 128,000 tokens, while certain Batch API configurations can reach 300,000 output tokens.

Pricing and usage costs

On the Claude API, Anthropic lists the primary token prices as:

  • Input: $2 per million tokens
  • Output: $10 per million tokens

Input tokens are the text, images, and other request content sent to the model. Output tokens are the generated response. The difference matters for applications that repeatedly send large prompts or produce long responses: a workflow that generates extensive code or documents can incur substantially more output cost than a short-answer application.

Claude Sonnet 5.5 also supports prompt caching. A five-minute cache write costs $2.50 per million tokens, while a one-hour cache write costs $4 per million tokens. Cache reads cost $0.20 per million tokens. The minimum cacheable prompt length is 512 tokens. Caching can reduce the cost of repeatedly sending the same large instructions, reference documents, or codebase context, although developers need to account for the separate write and read pricing.

The Batch API applies a 50% discount to input and output pricing. Batch processing is intended for requests that do not require immediate responses, such as bulk classification, document processing, offline code analysis, or large-scale content transformation. With the documented beta configuration, batch responses can also use the higher 300,000-token output limit.

Inputs, outputs, and reasoning

Sonnet 5.5 accepts text and images and returns text. Its image capability is for understanding visual content rather than generating images. For example, an application can use it to interpret a chart, inspect a screenshot, analyze a scanned document, or discuss an image alongside written instructions.

The model supports adaptive thinking. In practical terms, it can allocate more or less internal reasoning effort according to the task and the configured effort level. This is useful when a simple request should receive a quick answer but a difficult coding, planning, or analysis task needs more deliberate processing. Thinking behavior is not identical to ordinary visible prose: applications migrating from earlier Claude versions should review how thinking blocks are preserved, displayed, and handled between tool calls.

Anthropic’s documentation identifies structured-output workflows as supported. The supplied research does not establish a separate, general-purpose JSON mode, so developers should distinguish structured outputs from a dedicated JSON-mode setting when designing an integration.

Coding, tools, and workflow support

Coding is one of Sonnet 5.5’s main intended uses. Anthropic describes it as suitable for agentic coding, bug fixing, code review, implementation work, and maintaining codebases over longer tasks. The large context window can help an application provide more files, documentation, or repository history in a single interaction, while tool support lets the model participate in workflows that require actions outside text generation.

The model supports server-side and client-side tools, streaming, prompt caching, the Files API, PDF-oriented workflows, and batch processing. Tool use does not mean that the model independently has unrestricted access to a computer or external systems. The application still needs to define available tools, permissions, validation, and error handling.

Computer-use integrations require particular attention. On the Claude API and Google Cloud, Sonnet 5.5 supports the newer computer_toolset_20260801 toolset. The older computer_20251124 tool is rejected on those platforms. This is a concrete migration issue for applications that previously used an earlier Claude model or computer-use interface.

Sonnet 5.5 also rejects non-default values for temperature, top_p, and top_k. Applications that attempt to carry over tuning parameters from another model may therefore receive errors rather than merely different output. Developers should check the model-specific migration documentation before changing production traffic.

Strengths and trade-offs

The main strength of Claude Sonnet 5.5 is its balance. It is designed to be faster and cheaper than Claude Opus 5.5 while offering more capability than a lightweight model is intended to provide. That balance is particularly relevant for coding agents and business workflows where every request does not justify the highest available reasoning tier.

  • Long-context work: The 1 million-token context window can accommodate large document sets, extensive code context, and long conversations.
  • Software development: Coding, debugging, code review, and codebase maintenance are central use cases.
  • Document workflows: The model can support writing, document creation, spreadsheet and presentation tasks, PDF processing, and file-based analysis.
  • Visual understanding: Text-and-image input enables analysis of charts, screenshots, and other visual material.
  • Agentic applications: Tools, streaming, adaptive thinking, and batch processing support applications that go beyond one-off chat responses.
  • Operational cost: Its listed input and output prices are below the cost positioning expected of Anthropic’s Opus tier, while batch discounts and caching can reduce some workload costs further.

These advantages involve trade-offs. A 1 million-token context does not guarantee that every detail in a very large prompt will receive equal attention, and long outputs can still be expensive at the listed output rate. Adaptive thinking can also introduce different response behavior and integration requirements than older fixed-thinking configurations.

Limitations and compatibility considerations

Claude Sonnet 5.5 produces text only. It does not natively generate images, audio, video, music, or embeddings. Applications that require those output types need a separate model or service. Its image support should therefore be understood as visual input and image understanding, not image creation.

Sonnet 5.5 is not necessarily the best choice for every difficult task. Anthropic positions Claude Opus 5.5 above it for the most demanding reasoning and judgment-heavy work. Conversely, a lower-cost Haiku model may be more appropriate when response quality requirements are modest and request volume or latency dominates the decision.

Migration from Claude Sonnet 5 requires more than changing the model identifier. Developers should review adaptive-thinking defaults, the minimum thinking setting, preserved thinking blocks, computer-use tool types, advisor compatibility, and the restrictions on sampling parameters. If an application displays text between tool calls, it should also account for the possibility that this text is returned inside thinking blocks and configure its display behavior accordingly.

When to choose Claude Sonnet 5.5

Choose Claude Sonnet 5.5 when a project needs strong coding and knowledge-work performance but also needs fast responses and a lower operating cost than a top-tier reasoning model. It is a good fit for coding agents, code review systems, long-context document analysis, professional writing tools, image-aware assistants, and business automation that calls external tools.

It is especially suitable when the same application must handle varied workloads: a short technical question, a large repository review, a document-generation task, or an image-and-text analysis request. Prompt caching can help when those workflows repeatedly reuse a long system prompt or reference collection, and the Batch API is useful for high-volume work that can run asynchronously.

Consider Claude Opus 5.5 instead when the task is unusually judgment-heavy or when maximum capability is more important than speed and cost. Consider a Haiku-tier option when the work is simple, highly repetitive, or cost-sensitive enough that Sonnet-level capability is unnecessary. Choose another model or service when the application must generate media, create embeddings, or produce native audio or video.

Overall, Claude Sonnet 5.5 is best understood as Anthropic’s fast middle-to-upper-tier model: more capable and context-rich than a lightweight option, but positioned below Opus for the hardest work. Its value depends on using the capabilities that distinguish it—long context, coding, adaptive thinking, visual understanding, and tools—while designing around its text-only output and model-specific integration rules.


Answers to Frequently Asked Questions

What should developers consider when migrating to Claude Sonnet 5.5?
Migration requires more than changing the model identifier. Developers should review adaptive-thinking behavior, preserved thinking blocks, computer-use compatibility, advisor support, and sampling restrictions. On the Claude API and Google Cloud, the newer computer-use toolset is `computer_toolset_20260801`; the older `computer_20251124` tool is rejected. Non-default values for `temperature`, `top_p`, and `top_k` are also rejected.
What tools and input types does Claude Sonnet 5.5 support?
Claude Sonnet 5.5 accepts text and images and produces text. It supports adaptive thinking, server-side and client-side tools, streaming, prompt caching, the Files API, PDF-oriented workflows, and batch processing. Applications must define tool permissions and external actions because the model does not have unrestricted system access.
How much does Claude Sonnet 5.5 cost?
On the Claude API, Claude Sonnet 5.5 costs $2 per million input tokens and $10 per million output tokens. Prompt caching, which has separate write and read costs, is supported, and the Batch API provides a 50% discount on input and output pricing.
What is Claude Sonnet 5.5 designed for?
Claude Sonnet 5.5 is Anthropic’s general-purpose model for coding, codebase maintenance, long-context document analysis, professional writing, spreadsheets, presentations, image understanding, and tool-based workflows. It is positioned between Claude Opus 5.5 and the lower-cost Haiku line.
What is the context window and maximum output size of Claude Sonnet 5.5?
Claude Sonnet 5.5 has a 1,000,000-token context window. Its standard maximum output is 128,000 tokens, while supported Batch API configurations can generate up to 300,000 output tokens.


Sources 4
Provider

About Claude